ExplorerRoboticsRobotics
Research PaperResearchia:202606.26011

RouterVLA: Turning Smoke Tests into Supervision for Heterogeneous VLA Selection

Xingyu Ren

Abstract

We study whether pre-deployment evaluation rollouts can be reused to supervise policy selection. Robot teams routinely smoke test candidate vision-language-action (VLA) policies, then compress those trials into a global winner. RouterVLA evaluates this idea with outcome-disjoint cross-fitting: recorded probes build a profile for each frozen expert, and a separate trial scores the selected expert without entering its profile. Across 34,752 LIBERO-Plus rollout records, a transparent probe-success ...

Submitted: June 26, 2026Subjects: Robotics; Robotics

Description / Details

We study whether pre-deployment evaluation rollouts can be reused to supervise policy selection. Robot teams routinely smoke test candidate vision-language-action (VLA) policies, then compress those trials into a global winner. RouterVLA evaluates this idea with outcome-disjoint cross-fitting: recorded probes build a profile for each frozen expert, and a separate trial scores the selected expert without entering its profile. Across 34,752 LIBERO-Plus rollout records, a transparent probe-success rule raises held-out success from 0.4686 to 0.6149, a +14.64pp gain. Under the scalar-only profiles studied here, learned scorers are statistically indistinguishable from this rule, showing that commissioning carries the routing value while extra scalar scorer capacity does not create it. Reusing the scored trial inflates the measured gain by 1.87×1.87\times, so credible ledger routing needs outcome separation; model scaling improves individual policies, while commissioning-aware routing improves the system built from them.


Source: arXiv:2606.27355v1 - http://arxiv.org/abs/2606.27355v1 PDF: https://arxiv.org/pdf/2606.27355v1 Original Link: http://arxiv.org/abs/2606.27355v1

Please sign in to join the discussion.

No comments yet. Be the first to share your thoughts!

Access Paper
View Source PDF
Submission Info
Date:
Jun 26, 2026
Topic:
Robotics
Area:
Robotics
Comments:
0
Bookmark
RouterVLA: Turning Smoke Tests into Supervision for Heterogeneous VLA Selection | Researchia