OpenBot
Egocentric datasetGated
OBRS 29Bronze

Wearable AI

A benchmark for long-form QA, conversational QA, and proactive assistant behavior over first-person wearable-camera videos.

Scale
2,100 clips
Formats
JSONL · MP4
License
MIT
Published
2026-05-15

Decision summary

Best for

Agent evaluation over streaming first-person video

Main blocker

Less robot-action focused, but useful for evaluating an agent's temporal grounding.

Next check

Read the dataset manifest and feature schema.

Release history

Source-backed release timing for this canonical dataset record.

Important releases
  1. Dataset series published

    Release timing is recorded from the official source creation date.

    Release evidence

Selection readiness

6 evidence dimensions for deciding whether this dataset is ready to inspect, compare, or adopt. This is not a model benchmark.

Catalog evidence · not task performance

69/100

Provisional

3/6 dimensions scored · 41% confidence

Access and governance75conf. 85
Schema and signal coverageNot scoredconf. 15
Policy training readinessNot scoredconf. 15
World-model readiness68conf. 50
Failure and recovery readinessNot scoredconf. 15
Download and processing readiness65conf. 65
Read the evidence behind all 6 dimensions

Access and governance

Access and license are declared by the source.

usefulfit 75 · confidence 85

Schema and signal coverage

No machine-readable schema has been verified yet.

unknownnot scored · confidence 15

Policy training readiness

Observation/action alignment has not been established.

unknownnot scored · confidence 15

World-model readiness

Temporal observations plus geometry or semantic context are declared; sample alignment remains to be audited.

usefulfit 68 · confidence 50

Failure and recovery readiness

No verified failure/recovery annotation evidence is available yet.

unknownnot scored · confidence 15

Download and processing readiness

Scale is declared; transfer and processing estimates are not measured.

usefulfit 65 · confidence 65
Review unresolved evidence and next checks

Signal gaps

  • Gaze / attention. Not enough evidence is available to classify this signal.
  • Sim-real pairing. Not enough evidence is available to classify this signal.

Next checks

  • Read the dataset manifest and feature schema.
  • Run a bounded sample audit before assigning Strong readiness.

Dataset facts

Source
Hugging Face · facebook/wearable-ai
Evidence
official claim
Formats
JSONL · MP4 · starter kit
clips
2,100
tasks
3
size
~317 GB

Loop signals

5/7 present or partial

Gaze / attention

No decision-grade evidence captured yet.

unknown
Sim-real pairing

No decision-grade evidence captured yet.

unknown
View 5 more signal categories
Observation / ego video

video · Agent evaluation over streaming first-person video · Conversation-grounded video retrieval · A benchmark for long-form QA, conversational QA, and proactive assistant behavior over first-person wearable-camera videos.

present
Action / hand pose / robot state

Less robot-action focused, but useful for evaluating an agent's temporal grounding.

partial
Language intent / task phase

language · task phase

present
Feedback / correction / failure

Agent evaluation over streaming first-person video

present
License / format / access

Gated · MIT · JSONL · MP4

present
Evidence details and provenance

Official signal claims

VideoLanguageLanguageLanguageTask phase

Schema and annotations

No machine-readable schema facts are captured.

Sample verification

Pending. Metadata does not prove sample coverage, alignment, or file integrity.

curated source metadata

Integration notes

  • Less robot-action focused, but useful for evaluating an agent's temporal grounding.
  • Access requires accepting dataset terms on Hugging Face.