OpenBot
VLAClosed / referenced

RT-2

Google DeepMind

Vision-language-action model that transfers web-scale vision-language knowledge into robotic control.

Code
Not listed
Weights
Not listed
Checkpoint
Not listed
License
Unverified

Decision summary

Best for

Semantic generalization

Main blocker

Hardware, latency, dependencies, and checkpoint loading are not pipeline-tested.

Next check

Verify repository and artifact licenses.

Release history

Source-backed release timing for this canonical model record.

Important releases
  1. Model series introduced

    Release timing is anchored to the cited paper publication date.

    Paper

Evidence profile

A visual read of adoption evidence. Scores describe catalog evidence readiness, not task performance.

Published evaluation

1/6

dimensions scored

Unknown evidence remains visible and is never treated as a zero.

Access and governanceNot scoredconf. 20
Artifact availabilityNot scoredconf. 15
Training and loading reproducibilityNot scoredconf. 15
Training data requirements76conf. 70
Evaluation evidenceNot scoredconf. 10
Deployment readinessNot scoredconf. 10

Why these scores

The strongest decision reasons behind the evidence profile.

Top 3 adoption signals

Artifact availability

Unknown

Code and weights are not verified.

· confidence 15

Training and loading reproducibility

Unknown

No verified loading configuration is available.

· confidence 15

Evaluation evidence

Unknown

No structured evaluation evidence has been verified.

· confidence 10

Data requirements

Declared loop-data needs, missing evidence, and the next checks that matter.

4 required categories

Observation / ego video

observation

required

Language intent / task phase

language intent

required

Action / robot state

actions

required

Feedback / correction / failure

task success

required

Critical gaps

Core categories are represented. Interface alignment and data quality still require verification.

Linked datasets

Signal-level links only. Verify runtime interfaces before use.

Browse datasets
More evidenceShow

Additional data signals

Sim-real / embodiment metadata

Descriptive model record

useful

Additional decision dimensions

Access and governanceUnknown

License or access terms are not verified.

· c20
Training data requirementsUseful

Required signal categories are structured; exact tensor and action interfaces still need verification.

76 · c70
Deployment readinessUnknown

Hardware, latency, dependencies, and checkpoint loading are not pipeline-tested.

· c10

Artifact facts

release_year
official_claim
2023
curated official source year
model.required_signals
official_claim
observation · language intent · actions · object labels · task success
Official model documentation and curated signal mapping
release_timing
metadata_verified
2023-07-28T21:18:02.000Z
arXiv 2307.15818 published
discovery.source
secondary_claim
name: worldbench/awesome-embodied-data-pyramid · trust: discovery_only · revision: 0568599d0619f20946e8089f6a941d0d9e30b690
#table-embodied-foundation-models-vla-wam, table 17, row 2
discovery.raw_fields
secondary_claim
Time: 2023.7 · Method: RT-2 · Institution: DeepMind · Project: [![Website](https://img.shields.io/badge/Website-0A7DBD)](https://robotics-transformer2.github.io/) · Model: VLA · Data: ![Real][data-real] ![General][data-general]
#table-embodied-foundation-models-vla-wam, table 17, row 2
model.reported_release_time
secondary_claim
2023.7
#table-embodied-foundation-models-vla-wam, table 17, row 2
model.type
secondary_claim
VLA
#table-embodied-foundation-models-vla-wam, table 17, row 2
model.training_data_layers
secondary_claim
real_robot · general
#table-embodied-foundation-models-vla-wam, table 17, row 2
model.institution
secondary_claim
DeepMind
#table-embodied-foundation-models-vla-wam, table 17, row 2
catalog.curation
secondary_claim
tier: editorial_focus · collection: WorldBench Awesome Embodied Data Pyramid · policy: human_curated_priority
#table-embodied-foundation-models-vla-wam, table 17, row 2

Additional references

OpenBot notes

  • Useful reference point for VLA direction even if the model itself is not open.
  • Highlights why OpenBot should track both language/task labels and action traces.