OpenBot
Back to datasets
Egocentric datasetOpenReadiness 74 · confidence 34

POV Egocentric Video Robotics FHD Samples Dataset

MIT-licensed POV robotics video samples

<1Krows

A small MIT-licensed first-person robotics video sample set for human demonstrations, VLA experiments, manipulation, and household activities.

Best for

Small open fixture for first-person video ingestion

Not for / blocker

Small sample dataset, not a large training corpus.

Download decision

Inspect schema and run a bounded sample audit before committing to the full release.

Policy learningUseful

Has observation, action/state proxy, and task or language context.

fit 85 · confidence 55

World modelUseful

Has rich observation and semantic context, but limited geometry/sim-real alignment.

fit 68 · confidence 55

WAMUseful

Contains observation, intent, and action/state, but feedback/correction signal is weak.

fit 72 · confidence 55

Verified facts and provenance

Claims, metadata verification, and sample verification are shown separately.

curated official source
Official claim · signals
VideoTask phaseDemonstrationsRobotics Labels
Metadata verified · schema / annotations

Unknown — no machine-readable schema facts have been captured.

Sample / pipeline verification

Unknown — metadata conclusions do not prove sample coverage, alignment, or file integrity.

Declared loop signal coverage

Signals inferred from official metadata; Data pipeline verification is still pending.

4/7 categories present or partial

Observation / ego video

video · Small open fixture for first-person video ingestion · MIT-licensed POV robotics video samples · A small MIT-licensed first-person robotics video sample set for human demonstrations, VLA experiments, manipulation, and household activities.

present

Action / hand pose / robot state

Household manipulation demos · A small MIT-licensed first-person robotics video sample set for human demonstrations, VLA experiments, manipulation, and household activities.

partial

Gaze / attention

No decision-grade evidence captured yet.

unknown

Language intent / task phase

task phase · robotics labels

present

Feedback / correction / failure

No decision-grade evidence captured yet.

unknown

Sim-real pairing

No decision-grade evidence captured yet.

unknown

License / format / access

Open · MIT · video · text metadata

present

Model and task fit · OpenBot inference

Policy learningUseful

Has observation, action/state proxy, and task or language context.

fit 85 · confidence 55

World modelUseful

Has rich observation and semantic context, but limited geometry/sim-real alignment.

fit 68 · confidence 55

WAMUseful

Contains observation, intent, and action/state, but feedback/correction signal is weak.

fit 72 · confidence 55

Failure miningUnknown

Failure, correction, intervention, and recovery annotations have not been verified.

fit 50 · confidence 15

Good tasks

household manipulationpick-place / manipulationvideo-language reasoning

Blockers and unresolved evidence

  • Gaze / attentionunknown
    Not enough evidence to classify this signal. Verify metadata or a bounded sample.
  • Feedback / correction / failureunknown
    Not enough evidence to classify this signal. Verify metadata or a bounded sample.
  • Sim-real pairingunknown
    Not enough evidence to classify this signal. Verify metadata or a bounded sample.

Raw dataset signals

VideoTask phaseDemonstrationsRobotics Labels

OpenBot fit

  • Small open fixture for first-person video ingestion
  • Household manipulation demos
  • VLA data catalog examples

Integration notes

  • Small sample dataset, not a large training corpus.
  • Useful because it is MIT licensed and directly accessible.

Related by signals