Can robots stop getting distracted by visual clutter and actually focus on what matters for the task? Researchers from Fudan, SJTU, and HKU present GuidedVLA. Instead of treating the action decoder as a single black box, they split it into specialized “attention heads”—each
