VLA models
SmolVLA
A 450M open VLA trained only on crowd-sourced community data, small enough to train on a single consumer GPU and run on CPU or a MacBook; asynchronous inference yields ~30% faster response and ~2x task throughput. Built by Hugging Face's Paris-based LeRobot team.
Sources : arXiv (2025-06-02)Hugging Face (2025-06-03)Hugging Face (model card) (2026-07-09)
← Back to comparator · VLA models