RoboAtlas

Robot VLA Models

Explore vision-language-action models for robotics. Compare published versions, checkpoints, robot compatibility, licenses and benchmark evidence.

Definition

Vision-language-action (VLA) models use visual observations and language instructions to produce actions for robots.

Inclusion criteria

Includes published model families with a released version classified as VLA in the RoboAtlas model catalog. Other embodied AI categories are listed separately.

Key fields

Compare model versions, action spaces, checkpoints, licenses, robot compatibility and benchmark evidence on the linked model profiles.

Data gaps

A VLA label does not establish general-purpose capability. Training data, hardware requirements and evaluation conditions may be undisclosed.

Current coverage

13 published VLA model families

Catalog last updated: Sep 28, 2026

VLA model examples

How to compare

Open the VLA catalog, select released versions and compare their documented capabilities and sources.

Browse VLA models
Submit a source or correction