Back
精读 LingBot-VA 2.0:以语义视觉-动作 tokenizer、因果 MoE DiT、Multi-Chunk Prediction 与异步闭环推理,实现可泛化的 video-action robot control。
paper deep dive
lingbot-va
world action model
robot learning
mixture of experts
精读 LingBot-VA:将视频世界建模与逆动力学动作生成交织为因果自回归序列,并以 MoT、KV Cache、部分去噪和 FDM 校正实现长时程闭环操控。
world model
vla
robot control
flow matching