arXivRobotics
"Rephrase Before You Act": Mitigating the Extreme Language Sensitivity of Vision-Language-Action Models
機器人控制的「文字敏感症」:為何一個詞能讓 VLA 模型成功率從 100% 跌到 2%?
This study exposes the extreme sensitivity of Vision-Language-Action (VLA) models to instruction phrasing and introduces a zero-shot framework that uses LLMs to distill rewriting rules, significantly boosting robotic task success without retraining.
2 min read