What if building a robot that sees, understands, and acts was as easy as snapping Lego together? Enter StarVLA: a modular codebase that lets you swap vision-language or world-model backbones and action heads independently. It matches or surpasses prior methods on benchmarks
StarVLA: Modular Vision-Language Robot Codebase
By
–
