/
← Accept All   Archive
Apple Machine Learning ResearchResearch

REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs

September 2

Most current vision-language-action (VLA) models—such as OpenVLA, π0, RT-2, and RDT-1B—are “monolithic.” This means they generate raw motor commands or very short sequences of actions, without organizing behaviors into r

Read at Apple Machine Learning Research ↗

More from Apple Machine Learning Research on Accept All.