A low-cost physical regression station for VLA policy releases
VLA teams can add a small real-robot station to their release process and use it as a recurring regression test. VLA-REPLICA gives a concrete bill of materials and protocol: an SO-101 6-DoF arm, RealSense D455 top camera, wrist webcam, light box, commodity objects, AprilTag alignment, fixed lighting, reference images, and predefined object placements. The full setup is reported at about $1050, and one user built it in under an hour.
The useful adoption change is a fixed local gate before claiming progress on a new policy or fine-tune. Run the same 10 manipulation tasks, reuse the 500-demonstration adaptation set, and report success across the 50 in-distribution and 40 out-of-distribution scenes. The baseline table also sets expectations: π0.5 leads the tested policies at 0.54 average success in-distribution, while ACT, DiT variants, SmolVLA, X-VLA, and π0 are lower. A team with a new VLA can learn quickly whether a leaderboard gain still works with lighting, camera pose, object placement, and contact on real hardware.