Back
精读 FAST:用 DCT 与 BPE 压缩高频机器人动作,把自回归 VLA 从“预测数百个冗余 token”变成高信息密度的动作序列。
paper deep dive
fast
action tokenization
vision-language-action model