VLA models

RFM-1

United States

Early commercial robotics foundation model for warehouse manipulation, including a learned physics world model that predicts video outcomes of candidate actions. In August 2024 Amazon hired Covariant's founders and about a quarter of its staff and took a non-exclusive license to its models.

Updated 2026-07-09 · Last verified 2026-07-09

OrganizationCovariant
CountryUS
Release2024-03
Params (B)8
Open weightsno
Architecture8B any-to-any multimodal transformer performing autoregressive next-token prediction across text, images, video, robot actions and numerical sensor readings
Modalitiesvision, language, action, video, sensor
Embodimentsmanipulator
Training dataInternet data plus Covariant's proprietary multimodal warehouse production data (Covariant Brain fleet: picking, sortation, induction, depalletization since 2017)

Sources : Covariant (2024-03-11)IEEE Spectrum (2024-03-11)TechCrunch (2024-08-31)

← Back to comparator · VLA models
Christian Verbrugge

Christian Verbrugge · D·Fairy

Bringing a physical AI solution into European industry?

15 years at KUKA on automotive OEM programs (€100M), now focused on edge AI and autonomous agents. I support robotics and embodied AI companies on their EMEA market entry: OEM and Tier 1 access, co-funded pilots, AI Act readiness.

Book an intro call