MEMS Robotics Seminar: How general are generalist robot policies? Data scaling, diagnostic tools, and memorization in VLAs
Abstract: Vision-Language-Action (VLA) policies have recently emerged as a promising paradigm for generalist robot autonomy. However, VLAs have several challenges that must be overcome before they can achieve their potential. Firstly, these models require fine-tuning with human-teleoperation demonstrations, which can be tedious, expensive, and time-consuming to collect. Secondly, policy performance is limited to teleop demonstration […]