Behavioral phasing for precise manipulation
Move-Then-Operate argues that high-precision manipulation improves when the policy treats approach motion and contact work as separate behaviors. Its dual-expert design uses one expert for moving and one for operating, with a router selecting the phase per action chunk. On 8 RoboTwin2 tasks with 50 demonstrations per task, it reports 68.88% average success, above pi_0 at 44.75%, RDT at 35.63%, and ACT at 31.63%. The gains are large on contact-heavy tasks such as Click Bell at 99% versus 44% for pi_0 and Place Cans Plasticbox at 79% versus 34%. The paper also claims peak performance in 40% fewer training steps.