RoboIRS: Inference-Time Internal Representation Steering for Generalist Robot Policies

📅 2026-10-03
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the performance degradation of vision-language-action models on out-of-distribution tasks by proposing an inference-time internal representation steering method. Without updating policy parameters, the approach trains a linear classifier solely on successful and failed trajectories to identify intervention points and derive steering directions, enabling effective intervention on both the frozen policy π0.5 and the world action model Cosmos Policy. Experimental results demonstrate that the success rate in simulated tasks improves from 44.4% to 66.2%, while real-world robotic manipulation performance is also substantially enhanced. These findings validate the effectiveness and generality of this retraining-free paradigm in restoring model generalization capabilities.
📝 Abstract
Vision-language-action (VLA) and world-action models (WAMs) often degrade under out-of-distribution task variations despite retaining partial task capability. To recover such capability, we propose RoboIRS, an inference-time internal representation steering method that uses successful and failed rollouts to train linear classifiers, select outcome-relevant intervention locations, and derive task-specific steering directions without updating policy parameters. On 15 simulation tasks with a frozen $\pi 0.5$ policy, RoboIRS improves the average success rate from 44.4% to 66.2%, outperforming alternative inference-time intervention baselines while adding little inference time. We further validate RoboIRS on real-robot manipulation using the same $\pi 0.5$ policy and demonstrate its applicability to a world-action model Cosmos Policy, where the average success rate improves from 35.4% to 55.4%. These results show that directly steering internal robot-policy representations can improve the performance of robot policies at inference time. Project website is available at https://rollingoat.github.io/roboirs/.
Problem

Research questions and friction points this paper is trying to address.

Vision-Language-Action Models
World-Action Models
Out-of-Distribution
Robot Policies
Inference-Time Intervention
Innovation

Methods, ideas, or system contributions that make the work stand out.

Inference-time intervention
Internal representation steering
Vision-language-action models
World-action models
Generalist robot policies
💼 Related Jobs
No related jobs found.
J
Jiuzhou Lei
J. Mike Walker ’66 Department of Mechanical Engineering, Texas A&M University, College Station, TX 77843, USA
C
Chang Liu
J. Mike Walker ’66 Department of Mechanical Engineering, Texas A&M University, College Station, TX 77843, USA
D
Dayou Li
J. Mike Walker ’66 Department of Mechanical Engineering, Texas A&M University, College Station, TX 77843, USA
Z
Zhiyuan Zhang
Edwardson School of Industrial Engineering, Purdue University, West Lafayette, IN, USA
Xiao Liang
Xiao Liang
Zachry Department of Civil & Environmental Engineering, Texas A&M University
Adaptive RoboticsInfrastructure InspectionStructural MonitoringRobotic Disassembly
Yu She
Yu She
Assistant Professor, Purdue University
Robotic ManipulationMechanism DesignTactile SensingRobot Learning
Z
Zhiwen Fan
Department of Electrical and Computer Engineering, Texas A&M University, College Station, TX 77843, USA
Minghui Zheng
Minghui Zheng
J. Mike Walker '66 Department of Mechanical Engineering, Texas A&M University
RoboticsPlanningControlRobotic DisassemblyRemanufacturing Automation