Robotics: Science and Systems XXII

π*₀.₆: a VLA That Learns From Experience

Ali Amin, Raichelle Aniceto, Ashwin Balakrishna, Kevin Black, Ken Conley, Grace B. Connors, James Darpinian, Karan Dhabalia, Jared Di Carlo, Danny Driess, Michael Robert Equi, Adnan Esmail, Yunhao Fang, Chelsea Finn, Catherine Glossop, Thomas Godden, Ivan Goryachev, Lachy Groom, Hunter Hancock, Karol Hausman, Gashon Hussein, Brian Ichter, Szymon Jakubczak, Rowan Jen, Tim Jones, Benjamin Katz, Liyiming Ke, Chandra Kuchi, Marinda Lamb, Devin Leblanc, Sergey Levine, Adrian Li-Bell, Yao Lu, Vishnu Mano, Mohith Mothukuri, Suraj Nair, Karl Pertsch, Allen Z. Ren, Charvi Sharma, Lucy Xiaoyang Shi, Laura Smith, Jost Tobias Springenberg, Kyle Stachowicz, Will Stoeckle, Alexander Swerdlow, James Tanner, Marcel Torne, Quan Vuong, Anna Walling, Haohuan Wang, Blake Williams, Sukwon Yoo, Lili Yu, Ury Zhilinsky, Zhiyuan Zhou

Abstract:

Vision–language–action (VLA) models offer a promising path toward general-purpose robots, but achieving the reliability and speed required for practical deployment remains challenging. We present a general-purpose method, RL with Experience and Corrections via Advantage-conditioned Policies (RECAP) that improves the efficiency and reliability of VLA policies by utilizing their real-world experience. Our method introduces value-based advantage conditioning during both pre-training and post-training phases, enabling VLA policies to ingest highly heterogeneous real-world experience, including human demonstrations, policy rollouts, and online correction data. We show that the π0.6* model, trained with RECAP, achieves hours-long deployment of folding diverse laundry in real homes, can reliably assemble boxes in a factory, and make espresso drinks using a professional espresso machine. On some of the hardest tasks, RECAP more than doubles task throughput and roughly halves the task failure rate.

Download:

Bibtex:

  
@INPROCEEDINGS{AminA-RSS-26, 
    AUTHOR    = {Ali Amin AND Raichelle Aniceto AND Ashwin Balakrishna AND Kevin Black AND Ken Conley AND Grace B. Connors AND James Darpinian AND Karan Dhabalia AND Jared Di Carlo AND Danny Driess AND Michael Robert Equi AND Adnan Esmail AND Yunhao Fang AND Chelsea Finn AND Catherine Glossop AND Thomas Godden AND Ivan Goryachev AND Lachy Groom AND Hunter Hancock AND Karol Hausman AND Gashon Hussein AND Brian Ichter AND Szymon Jakubczak AND Rowan Jen AND Tim Jones AND Benjamin Katz AND Liyiming Ke AND Chandra Kuchi AND Marinda Lamb AND Devin Leblanc AND Sergey Levine AND Adrian Li-Bell AND Yao Lu AND Vishnu Mano AND Mohith Mothukuri AND Suraj Nair AND Karl Pertsch AND Allen Z. Ren AND Charvi Sharma AND Lucy Xiaoyang Shi AND Laura Smith AND Jost Tobias Springenberg AND Kyle Stachowicz AND Will Stoeckle AND Alexander Swerdlow AND James Tanner AND Marcel Torne AND Quan Vuong AND Anna Walling AND Haohuan Wang AND Blake Williams AND Sukwon Yoo AND Lili Yu AND Ury Zhilinsky AND Zhiyuan Zhou}, 
    TITLE     = {{π*₀.₆: a VLA That Learns From Experience}}, 
    BOOKTITLE = {Proceedings of Robotics: Science and Systems}, 
    YEAR      = {2026}, 
    ADDRESS   = {Sydney, Australia}, 
    MONTH     = {July}, 
    DOI       = {10.15607/RSS.2026.XXII.087} 
}