Egocentric video + synchronized 6-axis IMU from a wristwatch camera —
the exact viewpoint robot wrist cameras consume. The dataset packs below are ready for robot hand learning.
Pick a task below to see its sample video, price, and how to buy.
Pack 01
towels · garments — folding, unfolding, smoothing
300 episodes · video + IMU · commercial, non-exclusive license
$390 / ₩490,000
View & order →Pack 02
lightbulbs · hooks · reaching and turning overhead
300 episodes · video + IMU · commercial, non-exclusive license
$390 / ₩490,000
View & order →Pack 03
shoelaces · zippers · buttons · cable ties
300 episodes · video + IMU · commercial, non-exclusive license
$390 / ₩490,000
View & order →Pack 04
jars · bottles · screw caps · snap lids · boxes
300 episodes · video + IMU · commercial, non-exclusive license
$390 / ₩490,000
View & order →Pack 05
bags · pouches · unwrapping · tearing open
300 episodes · video + IMU · commercial, non-exclusive license
$390 / ₩490,000
View & order →Pack 06
setting the table · clearing · putting away (no cooking)
300 episodes · video + IMU · commercial, non-exclusive license
$390 / ₩490,000
View & order →Pack 07
shelving · drawer organizing · object placement
300 episodes · video + IMU · commercial, non-exclusive license
$390 / ₩490,000
View & order →Made to order
Need a task that isn't in the catalog? Describe the manipulation you want and we build the dataset for you: scope and invoice within 48 h, collection in 2–3 days after payment, delivery within about a week — with the same consented participants, hardware, and format as every pack above.
from ₩1,000,000
/ 100 episodes
Base price — the final quote arrives by invoice and varies with task complexity.
Order a custom dataset →Modern manipulation policies don’t watch from a distance — they read cameras mounted on the robot’s own wrists. π0, for example, takes three camera streams as input, and two of them are wrist views.
Most human video datasets are recorded from the head. A head camera sees what a person looks at — not what the hand touches; the hand–object contact region is routinely occluded. We record from a wristwatch camera —
every episode is captured from the viewpoint a robot’s wrist camera will use, with a synchronized IMU carrying the wrist’s own motion.
Every number below is measured on our hardware, not estimated.
| View | Wrist-mounted, palm-facing — continuous view of the hand–object contact region |
| Video | 640×480 @ 30 fps H.264 · 20-second complete task episodes (start state → end state) |
| IMU | 6-axis (ST LSM6DSO), ~100 Hz effective, ±16 g / ±2000 dps |
| Sync | Per-clip self-describing header; IMU logger leads video by ~0.6 s with the offset documented per clip |
| Integrity | 0.00% malformed rows · single-sample accel outliers ~1.4% in raw (removed in processed) |
| Format | LeRobot v3.0 + raw mp4/csv + sha256 manifest |
| Consent | Participants consented, incl. commercial redistribution. Serial-only IDs — no PII in the dataset |
Yes. The standard license is commercial non-exclusive: train, fine-tune, evaluate, and own the resulting models. Redistribution of the raw data is not allowed.
Task-level metadata ships with every pack: task id, subject id, success flag, and per-clip sensor headers. We do not provide hand-pose or action labels.
Yes — see 'Made to order' below the catalog. You specify the manipulation task and environment constraints; we collect with consented participants on our hardware.
No. This dataset is collected in dedicated sessions with separately consented participants. It contains no patient data, no medication records, and no data from our healthcare products.
Not under the standard tier — a Commercial + Redistribution tier (×2–3) covers derived-dataset publishing. Reselling the original data is never allowed.