MA-VLA: Multi-Arm Vision-Language-Action Model for Collaboration and Compositional Generalization Paper • 2608.25864 • Published 8 days ago • 9
How Far are VLMs from Visual Spatial Intelligence? A Benchmark-Driven Perspective Paper • 2509.18905 • Published Sep 23, 2025 • 31