Egocentric/Exocentric Video - Object Destruction, Debris Manipulation

Egocentric & First-Person Data

Tags and Keywords

Egocentric

Exocentric

Video

Ego-exo

Multi-view

Robotics

Human-object

Manipulation

Embodied

Action

Physics

Debris

Cleanup

Gopro

Egocentric/Exocentric Video - Object Destruction, Debris Manipulation Dataset on Opendatabay data marketplace

£55,000

About

EgoExo-Wreck: Synchronized Egocentric/Exocentric Video — Object Destruction, Debris Manipulation & Cleanup

EgoExo-Wreck is a synchronized egocentric + exocentric (first-person + third-person) video dataset captured at a commercial rage room facility in Las Vegas, Nevada, produced specifically for AI training. It documents two rare, physics-rich activity domains: high-force human-object interaction — real participants destroying electronics, furniture, glassware, and ceramics with hand tools (bats, crowbars, hammers) — and debris cleanup and environment reset — trained staff sweeping, collecting fragments, bagging debris, and re-placing objects across repeated procedural cycles (destroy → clean → reset).
Its significance: paired ego-exo capture with commercial AI-training rights is scarce (most ego-exo corpora are research-licensed), and the content covers manipulation of broken, irregular, and deformable objects in unstructured environments — object classes largely absent from household-activity datasets — plus real-world impact and shattering dynamics valuable for physics and world-model research. Every participant has signed explicit AI-training and third-party-licensing consent with a session-to-signature audit trail. The corpus grows ~80 hours per month through ongoing directed capture.

Data Product Features

  • Paired video streams: head-mounted egocentric 4K (GoPro Hero 12, 3840×2160 @ 30/60 fps) + fixed-mount exocentric (2688×1520 or 1920×1080 @ 30 fps), per session
  • Per-clip synchronization data: ego/exo sync frame anchors, lag in seconds, confidence rating, and method (audio waveform correlation, 8 kHz mono RMS-delta envelope; ±~50 ms accuracy)
  • Machine-readable manifest (CSV): clip IDs, session IDs, capture dates, activity class, room/location, durations, synchronized-overlap runtime, device metadata, frame rates, file sizes
  • Object-detection auto-labels: 16-class in-domain detector (room object inventory including tools), with documented training methodology and per-class metrics
  • Narration layer (newer cleanup capture): staff natural-language narration recorded on a separate audio channel with timestamped transcripts
  • Multi-person sessions: 1–6 participants per destruction session, including multi-agent coordination footage

Distribution

  • Format: MP4 video (H.264/H.265); CSV manifest; label files (YOLO/COCO format); narration transcripts (text with timestamps)
  • Data Volume: ~150 hours of synchronized ego-exo video across paired clips (per-clip synchronized-overlap runtime documented in manifest); growing ~80 hours/month; typical clip length 6–31 minutes; ego files ~1.5–20 GB per clip, exo files ~0.4–1 GB per clip
  • Structure: session-based pairing — each record links one egocentric file to its matched exocentric file with frame-level sync anchors
  • Delivery: secure file transfer or cloud bucket; evaluation sample available under click-through agreement

Usage

This data product is ideal for a variety of applications:
  • Robot manipulation / imitation learning: demonstrations of grasping, sweeping, and handling irregular, deformable, and fragmented objects (broken glass, ceramic shards, mixed debris) in unstructured settings
  • Vision-language-action (VLA) model training: paired ego-exo demonstrations with natural-language narration for instruction-conditioned learning
  • World models / video generation: real impact dynamics, shattering, and object state transitions (intact → damaged → fragment) from synchronized viewpoints
  • Cross-view understanding: ego-exo correspondence, view-invariant action recognition, cross-view retrieval
  • Activity recognition and temporal segmentation: repeated procedural cycles with consistent scenes and natural variation
  • Physics and damage modeling: before/during/after sequences of object destruction for simulation ground truth and damage-assessment models

Coverage

  • Geographic Coverage: United States (single facility, Las Vegas, Nevada) — 4 destruction rooms plus warehouse/storage reset areas
  • Time Range: January 2026 – present (ongoing monthly capture)
  • Demographics: adult participants, 1–6 per session; staff-performed cleanup sequences; demographic attributes not annotated

License

Proprietary

AI Training Rights

Licensee is granted a non-exclusive, worldwide, and perpetual right to:
  • Use the Data Product to train, fine-tune, and evaluate machine learning models, including large language models.
  • Incorporate Data Product content into models and commercialize resulting model outputs.
  • Create derivative works (model weights, embeddings, etc.) for any lawful purpose.
Restrictions:
  • The Data Product itself may not be sold, redistributed, or shared outside of licensed usage.
  • Licensee must comply with all applicable laws, including data protection and privacy regulations.

Who Can Use It

  • Robotics and embodied-AI labs: training manipulation policies and VLA models on human demonstrations of debris handling and object interaction
  • World-model and video-generation teams: learning physical dynamics from real destruction and state-transition footage
  • Computer vision researchers: ego-exo benchmarks, cross-view correspondence, action recognition, temporal segmentation
  • Simulation, game, and VFX developers: real shatter and impact reference for physics simulation ground truth
  • Insurance and damage-assessment AI developers: before/after damage sequences for assessment-model training

Data Dictionary

Column NameData TypeDescriptionPossible Values/Notes
Record IDstringSequential record identifier001–012 (sample); zero-padded
Session IDstringCapture session identifierFormat S{YYYY}-{MM}
Clip IDstringUnique clip identifierFormat PM-{X|H}-YYYY-MM-NNN; X = destruction, H = cleanup
Capture DatedateDate of capture sessionISO 8601 (YYYY-MM-DD)
ActivitystringActivity class description"High-force object destruction" or "Debris cleanup / reset manipulation"
Activity CodestringMachine-filterable activity classDESTRUCT, CLEANUP
LocationstringCapture location within facilityRoom 1–4, Storage/Warehouse
Ego DurationtimeEgocentric clip runtimeHH:MM:SS
Exo DurationtimeExocentric clip runtimeHH:MM:SS
Synced OverlaptimeRuntime where both views exist (usable paired footage)HH:MM:SS
Synced Overlap (min)floatSame, in decimal minutes for aggregatione.g., 27.5
Ego DevicestringEgocentric camera modelGoPro Hero 12
Ego ResolutionstringEgocentric resolution3840×2160
Ego FOV ModestringEgocentric field-of-view settingDocumented per clip
Ego FPSfloatEgocentric frame rate29.97 or 59.94 (per-clip)
Ego File Size (GB)floatEgocentric file size~1.5–20 GB
Exo DevicestringExocentric camera modelFixed-mount
Exo ResolutionstringExocentric resolution2688×1520 or 1920×1080
Exo FPSfloatExocentric frame rate30 (per-clip verified)
Exo File Size (GB)floatExocentric file size~0.4–1.0 GB
Ego Sync FrameintegerEgo frame index aligned to Exo Sync Frame≥ 0
Exo Sync FrameintegerExo frame index aligned to Ego Sync Frame≥ 0
Sync Lag (s)floatStart-time offset between streamsNegative = ego started first
Sync ConfidencestringConfidence rating of sync estimateHigh, Medium
Sync MethodstringSynchronization techniqueAudio waveform correlation (8 kHz mono, 0.05 s RMS-delta envelope)
Ego FilestringEgocentric filenameMatches Clip ID convention
Exo FilestringExocentric filenameMatches Clip ID convention

Additional notes: Sessions follow repeated procedural cycles (destroy → clean → reset) in fixed rooms with rotating object inventories, making the dataset well suited to skill-learning and procedural benchmarks. Delivered audio contains no third-party copyrighted music (cleanup sessions are captured music-free; destruction-session audio is delivered stripped or source-separated — specify preference at licensing). Directed-capture programs are available: licensees may specify camera configurations, object inventories, and task protocols for monthly ongoing delivery. An evaluation sample (manifest + representative clips) is available under a click-through evaluation agreement.

Listing Stats

VIEWS

15

DELIVERY

CUSTOM, S3

LISTED

11/07/2026

UPDATED

11/07/2026

REGION

NORTH AMERICA

Universal Data Trust Rating UDTRTRUST

5 / 5

Loading...

£55,000

Download Dataset in Unknown Format