You set the task, environment, viewpoint, volume and quality bar. We handle sites, operators, rigs, consent, redaction and QC.
2 to 6 synchronized cameras on live worksites, calibrated per unit. Enough views to recover 3D hand pose and 6-DoF trajectories, not just RGB.
Every team that evaluates us pressure-tests the same three things. Here is where we stand on each.
Collections running now in Korea, the UAE and the United States. Everything above is a loop we are already operating, not a capability we would stand up for you.
Your robot will not spend its life in one kind of room. Pick the settings you need and we line up sites and operators in each of them.
Collection runs in Korea, the UAE and the United States under one agreement. Running the same task family in three regions is the cheapest way to find out whether a policy transfers or whether it learned one country.
Running six synchronized cameras in an occupied worksite is a different operational problem than running one, and it is the range we already work in. You pick the count when we scope the collection.
One RGB sensor, light enough that operators stop noticing it. That is why this is the setup we can run for full shifts across a lot of sites.
A calibrated pair on a fixed baseline. Use it when you need disparity, depth or 3D structure around the hands and the object.
A synchronized array covering the work volume from several angles. Enough overlap to triangulate 3D hand pose and 6-DoF object trajectories, and to keep the hands visible when one view gets occluded.
Camera count, array geometry, calibration, sync, resolution and frame rate are fixed per collection and documented with the data. Tell us what your model reads and we will confirm the setup before anyone starts recording. Delivery format is part of the spec. Name the schema you work in, LeRobot, RLDS or your own, and the export gets built into the collection.
Rigs, training and quality control are where a collection breaks. We run the whole loop in house, so a multi-camera collection does not add a vendor to your chain or a handoff to your timeline. The acceptance bar is agreed in writing before collection starts, and every step below runs against it.
Raw video with a per-clip record is the baseline. Task segmentation is available, and interaction labels, outcome states and any other fields get defined with you before collection starts.
A bespoke collection starts by finding sites, recruiting people and getting them scheduled. For us all three already exist. Miso runs a live service network across Korea, the UAE and the United States, and Miso Motion adds capture, consent, redaction and quality control on top of work that is already booked.
Four fields and we send real clips with the per-clip records and capture specs behind them.