Apple Workshop on Human-Centered Machine Learning 2024: Engineering Better UIs via Collaboration with Screen-Aware Foundation Models
AuthorsKevin Moran (University of Central Florida)
Apple Workshop on Human-Centered Machine Learning 2024: Engineering Better UIs via Collaboration with Screen-Aware Foundation Models
AuthorsKevin Moran (University of Central Florida)
Apple is presenting new research at the European Conference on Computer Vision (ECCV), which takes place in Malmö, Sweden, from September 8–12. We are proud to again sponsor the biennial conference, which brings together the scientific and industrial research communities around ML and computer vision. Below is an overview of Apple’s participation at ECCV 2026.
REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs
September 2, 2026research area Computer Vision, research area Methods and Algorithms
Most current vision-language-action (VLA) models—such as OpenVLA, π0, RT-2, and RDT-1B—are “monolithic.” This means they generate raw motor commands or very short sequences of actions, without organizing behaviors into reusable, well-defined abstractions. As a result, these models perform poorly on long-horizon (multi-step) tasks, and it’s difficult to interpret what they have learned. Existing approaches for discovering skills often avoid the…