EgoSchema Dataset
Video-language model evaluation
A diagnostic video-language benchmark derived from Ego4D, designed to test temporal and causal reasoning over long first-person videos.
Video-language model evaluation
Not a manipulation training set, but helpful for evaluating whether agents understand long first-person context.
Inspect schema and run a bounded sample audit before committing to the full release.
Observation/action alignment has not been established.
not scored · confidence 15
World-model observation, geometry, or temporal semantics are not verified.
not scored · confidence 15
No verified failure/recovery annotation evidence is available yet.
not scored · confidence 15
Verified facts and provenance
Claims, metadata verification, and sample verification are shown separately.
Unknown — no machine-readable schema facts have been captured.
Unknown — metadata conclusions do not prove sample coverage, alignment, or file integrity.
Declared loop signal coverage
Signals inferred from official metadata; Data pipeline verification is still pending.
Observation / ego video
video · Ego4D video references · Video-language model evaluation · Underlying video access follows Ego4D licensing.
Action / hand pose / robot state
Not a manipulation training set, but helpful for evaluating whether agents understand long first-person context.
Gaze / attention
No decision-grade evidence captured yet.
Language intent / task phase
language · temporal reasoning labels · Video-language model evaluation · A diagnostic video-language benchmark derived from Ego4D, designed to test temporal and causal reasoning over long first-person videos.
Feedback / correction / failure
Video-language model evaluation
Sim-real pairing
No decision-grade evidence captured yet.
License / format / access
License required · MIT · JSON · Ego4D video references
Catalog decision scorecard
Access and license are declared by the source.
fit 75 · confidence 85
No machine-readable schema has been verified yet.
not scored · confidence 15
Observation/action alignment has not been established.
not scored · confidence 15
World-model observation, geometry, or temporal semantics are not verified.
not scored · confidence 15
No verified failure/recovery annotation evidence is available yet.
not scored · confidence 15
Scale is declared; transfer and processing estimates are not measured.
fit 65 · confidence 65
Good tasks
Blockers and unresolved evidence
- Gaze / attentionunknownNot enough evidence to classify this signal. Verify metadata or a bounded sample.
- Sim-real pairingunknownNot enough evidence to classify this signal. Verify metadata or a bounded sample.
- Read the dataset manifest and feature schema.
- Run a bounded sample audit before assigning Strong readiness.
Raw dataset signals
OpenBot fit
- Video-language model evaluation
- Long-horizon plan checking
- Narrative consistency tests for agents
Integration notes
- Not a manipulation training set, but helpful for evaluating whether agents understand long first-person context.
- Underlying video access follows Ego4D licensing.
