Adapting Vision-Language-Action Models to Unknown Visual Disruptions During Execution
We introduce Self-supervised Adaptation from Leftover Trajectories (SALT), which uses the leftover trajectory, the unexecuted part of the previous action chunk, as self-supervision for test-time adaptation.
ProofPaper ↗
Key points
- Visual disruptions can arise while a robot is executing a task, leaving a vision-language-action (VLA) policy to respond without knowing the disruption type or timing.
- Because consecutive chunks overlap in time, the leftover provides a temporally aligned target for the current prediction over the same future control interval.
- At the onset of a visual shift, the leftover can retain a plan formed before the corruption, so updating the policy toward it anchors the adaptation across the shift (Transition Anchoring).
- On LIBERO-10, SALT increases average success across five persistent visual corruptions from 43.9% to 53.2% with SmolVLA and from 58.7% to 66.0% with GR00T N1.7, while largely preserving nominal performance.
Sources (1)
- [1]Adapting Vision-Language-Action Models to Unknown Visual Disruptions During ExecutionarXiv (AI, ML, NLP, CV, robotics, multi-agent) · Oct 6, 08:21 AM
We introduce Self-supervised Adaptation from Leftover Trajectories (SALT), which uses the leftover trajectory, the unexecuted part of the previous action chunk, as self-supervision for test-time adaptation.
Visual disruptions can arise while a robot is executing a task, leaving a vision-language-action (VLA) policy to respond without knowing the disruption type or timing.
Extractive summary: sentences quoted from the sources.
Before this
- Oct 6, 2026StairVLA: Stage-Aware Hierarchical Action Generation for Vision-Language-Action Models
- Oct 6, 2026SMART: Zero-Shot Sim-to-Real Articulated Object Manipulation via Large-Scale Synthetic Pretraining
- Oct 6, 2026CARE: Certifying Acceleration for Vision-Language-Action Inference
- Oct 2, 2026FastOPD: On-Policy Distillation for Lightweight VLA Deployment
- Jul 30, 2026Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration
- Jul 28, 2026Gemini Robotics 2 brings whole body intelligence to robots