Back
精读 OmniVLA-RL:用 MoT 三专家与 Block-wise Causal Attention 融合空间、语义和动作,并以 Flow-GSPO 稳定优化 Flow Matching 策略。
paper deep dive
vla
spatial intelligence
reinforcement learning