AI Research

Rephrase Before You Act: Characterizing and Mitigating Language Sensitivity in Vision-Language-Action Models

Medium Severity Global
Date Occurred Oct 07, 2026 17:57 UTC
Event Type AI Research
Source arXiv
Recorded Oct 08, 2026
Full Description

arXiv: Rephrase Before You Act: Characterizing and Mitigating Language Sensitivity in Vision-Language-Action Models Vision-language-action models (VLAs) are strikingly sensitive to instruction phrasing and do not inherit the language robustness of the vision-language models they are built on. A one-word edit can move success by tens of points: $π_{0.5}$ turns on a LIBERO stove 100% of the time for "switch on the stove" and 2% for "switch on the hot plate", and a $π_0$ checkpoint finetuned with rephrase augmentation still shows swings of up to 61 points. We characterize this sensitivity with statistically test

AI Intelligence Layer

AI Categories

ethics application
Event Metadata
  • ID #37605
  • Type AI Research
  • Region Global
  • Severity Medium
  • Indexed Oct 08, 2026