OpenBot
Back to Explore
Manipulation datasetGatedProvisional fit 77 · confidence 48

Xperience-10M Dataset

World model pretraining

10,000hours

A large egocentric multimodal dataset with synchronized video streams, audio, depth, poses, mocap, IMU, and hierarchical language annotations.

Best for

World model pretraining

Not for / blocker

Very large and controlled-access; the practical OpenBot path is metadata indexing plus targeted subset pulls.

Download decision

Inspect schema and run a bounded sample audit before committing to the full release.

Policy training readinessuseful

Observation and action/state signals are declared; alignment quality still depends on sample verification.

fit 85 · confidence 55

World-model readinessuseful

Temporal observations plus geometry or semantic context are declared; sample alignment remains to be audited.

fit 82 · confidence 50

Failure and recovery readinessunknown

No verified failure/recovery annotation evidence is available yet.

not scored · confidence 15

Verified facts and provenance

Claims, metadata verification, and sample verification are shown separately.

curated source metadata
Official claim · signals
VideoAudioDepthCamera poseHand poseIMULanguage
Metadata verified · schema / annotations

Unknown — no machine-readable schema facts have been captured.

Sample / pipeline verification

Unknown — metadata conclusions do not prove sample coverage, alignment, or file integrity.

Declared loop signal coverage

Signals inferred from official metadata; Data pipeline verification is still pending.

5/7 categories present or partial

Observation / ego video

video · depth · camera pose · A large egocentric multimodal dataset with synchronized video streams, audio, depth, poses, mocap, IMU, and hierarchical language annotations.

present

Action / hand pose / robot state

camera pose · hand pose · A large egocentric multimodal dataset with synchronized video streams, audio, depth, poses, mocap, IMU, and hierarchical language annotations.

partial

Gaze / attention

No decision-grade evidence captured yet.

unknown

Language intent / task phase

language · A large egocentric multimodal dataset with synchronized video streams, audio, depth, poses, mocap, IMU, and hierarchical language annotations.

present

Feedback / correction / failure

No decision-grade evidence captured yet.

unknown

Sim-real pairing

depth · camera pose · Real-to-sim and sim-to-real data alignment · A large egocentric multimodal dataset with synchronized video streams, audio, depth, poses, mocap, IMU, and hierarchical language annotations.

present

License / format / access

Gated · Apache-2.0 · Hugging Face dataset · multimodal episode files

present

Catalog decision scorecard

Access and governanceuseful

Access and license are declared by the source.

fit 75 · confidence 85

Schema and signal coverageunknown

No machine-readable schema has been verified yet.

not scored · confidence 15

Policy training readinessuseful

Observation and action/state signals are declared; alignment quality still depends on sample verification.

fit 85 · confidence 55

World-model readinessuseful

Temporal observations plus geometry or semantic context are declared; sample alignment remains to be audited.

fit 82 · confidence 50

Failure and recovery readinessunknown

No verified failure/recovery annotation evidence is available yet.

not scored · confidence 15

Download and processing readinessuseful

Scale is declared; transfer and processing estimates are not measured.

fit 65 · confidence 65

Good tasks

World model pretrainingReal-to-sim and sim-to-real data alignmentMultimodal episode quality checks

Blockers and unresolved evidence

  • Gaze / attentionunknown
    Not enough evidence to classify this signal. Verify metadata or a bounded sample.
  • Feedback / correction / failureunknown
    Not enough evidence to classify this signal. Verify metadata or a bounded sample.
  • Read the dataset manifest and feature schema.
  • Run a bounded sample audit before assigning Strong readiness.

Raw dataset signals

VideoAudioDepthCamera poseHand poseIMULanguage

OpenBot fit

  • World model pretraining
  • Real-to-sim and sim-to-real data alignment
  • Multimodal episode quality checks

Related models and papers

Model references linked to similar loop signals.

Integration notes

  • Very large and controlled-access; the practical OpenBot path is metadata indexing plus targeted subset pulls.
  • Useful as a reference for the signals OpenBot Data should preserve.

Related by signals