VLA models

RFM-1

United States

Early commercial robotics foundation model for warehouse manipulation, including a learned physics world model that predicts video outcomes of candidate actions. In August 2024 Amazon hired Covariant's founders and about a quarter of its staff and took a non-exclusive license to its models.

Updated 2026-09-06 · Last verified 2026-09-06

OrganizationCovariant
CountryUS
Release2024-03
Params (B)8
Open weightsno
Architecture8B any-to-any multimodal transformer performing autoregressive next-token prediction across text, images, video, robot actions and numerical sensor readings
Licensenon-exclusive license to Amazon to use Covariant’s robotic foundation models
Modalitiesvision, language, action, video, sensor
Embodimentsmanipulator
Training dataInternet data plus Covariant's proprietary multimodal warehouse production data (Covariant Brain fleet: picking, sortation, induction, depalletization since 2017)

Sources: Covariant (2024-03-11)IEEE Spectrum (2024-03-11)TechCrunch (2024-08-31)

← Back to comparator · VLA models
Christian Verbrugge

Christian Verbrugge · D·Fairy

Bringing a physical AI solution into European industry?

15 years at KUKA on automotive OEM programs (€100M), now focused on edge AI and autonomous agents. I support robotics and embodied AI companies on their EMEA market entry: OEM and Tier 1 access, co-funded pilots, AI Act readiness.

Book an intro call