PLATFORM
How Roborecs captures the data.
From a first-person human demonstration to a labelled, license-ready training episode: the pipeline, the eight capture channels, the Phase-2 fidelity facility, and the specialist models the corpus trains.
How it works
How a demonstration becomes training-ready data, the phase roadmap, and every channel a session records.
From human demonstration to robot training data.
Egocentric demonstrations
Trained operators record real two-handed tasks from a first-person view on a wearable head-and-wrist rig: synchronized RGB, depth, 3D hand pose and skeleton, motion, and audio. Scales linearly with trained operators, at a fraction of teleoperation’s cost per hour, no lab or robot needed to record.
HEAD + WRIST RIG · MULTI-VIEW RGB · SUB-MS SYNCAlign and annotate
Sub-millisecond timestamps align every channel; faster streams are downsampled and slower ones interpolated. Each episode is segmented into tasks and steps, with object and contact tags, consent, and provenance.
SUB-MS SYNC · TASK + STEP LABELS · PROVENANCELicense-ready episodes
Output: labelled training episodes in the format the ecosystem already uses, ready to drop into humanoid foundation-model pipelines.
GDPR · EU AI ACT · HDF5 / LeRobot v3THE ROADMAP · CAPTURE TO CORPUS TO MODELS
One data engine, built in three turns.
Egocentric capture
First-person human demonstrations, captured on a wearable head-and-wrist rig. The cheapest, most scalable way to produce the base of every humanoid model.
The wide base of the corpus. Where we start.
Multimodal fidelity
In-facility capture adds the signals a wearable rig leaves out: force, torque, contact, and eye-gaze. The fidelity that force-controlled, two-handed assembly demands.
The fidelity the base climbs toward.
Specialist models
The corpus compounds and stays ours. When it reaches scale, and with the ML research leadership we bring on for this phase, we train the vertical models it powers, compatible with NVIDIA Isaac GR00T N1.7, Physical Intelligence π0, and Hugging Face LeRobot.
The data moat becomes a model moat.
This is what robotics used to be.
Eight synchronized channels, merged into one labelled frame.
The multimodal supply that humanoid robots train on.
Roborecs is building the pipeline that captures these channels first-person, six from Phase 1 with force and gaze added in the Phase 2 facility, and fuses them into training frames in the format the industry already uses.
Capture specification: the target channels, their rates, and the fused output. Illustrative, not live or recorded sensor data.
Each channel captures at its native rate. Sub-millisecond timestamps align every sample. Faster channels are downsampled; slower ones interpolated. The output is a single multimodal frame, 30 times per second.
A Roborecs dataset is delivered as documented, training-ready episodes, not raw dumps. What is available depends on the capture phase.
- 7-camera RGB video
- Depth
- IMU at 500 Hz
- Spatial trajectories
- Audio
- Hand pose and skeleton
- Force-torque and tactile (Phase 2)
- Task and step segmentation
- Object and interaction tags (pick, place, insert, route)
- Contact and grasp events
- Per-episode metadata (task, environment, operator)
- Custom taxonomies for your model or benchmark
- LeRobot v3 / HDF5
- Dataset cards and manifests
- Consent and provenance artifact per episode
- EU-jurisdiction delivery
- Cloud bucket or encrypted transfer
Delivery spec for the capture program. Available modalities and annotations depend on capture phase.
The fidelity layer
The Sofia facility, and the force and touch a camera alone cannot carry.

Built for physical AI, in Europe.
A purpose-built data capture facility planned for Sofia Tech Park, 1,100 m² at launch scaling to 5,000 m². Direct access to STEM talent, energy infrastructure, and EU logistics. Targeted operational from Q3 2027.
Force-controlled, bimanual assembly, the Phase 2 fidelity catalog.
Phase 2 adds the contact-rich tasks that need force and two-handed coordination, beyond what first-person video alone carries. Robot OEM clients can commission task libraries across the categories below, with custom commissions for proprietary needs.
- →Connector mating & cable routing
- →Precision fastening with torque
- →PCB & component handling
- →Pick & place assembly
- →Quality inspection
- →Bin sorting & packing
- →Bimanual coordination
- →Tool manipulation
- →Fine alignment
- →Collaborative task handoff
- →Assistance & guiding
- →Safe proximity work
- →OEM-defined task specs
- →Bespoke capture sessions
- →Full IP assignment
Custom Commissions are available for OEM-specific task libraries. Full IP assignment. EU GDPR-compliant provenance from day one.
DISCUSS COMMISSION →An operator drives. The robot performs.
The Phase 2 fidelity layer, for the hardest contact-rich tasks. Move your cursor to steer the operator command; the robot executes while we record its joint states, forces, and contacts as it works, captured in the robot’s own action space, directly trainable, with no human-to-robot retargeting.
illustrative telemetry · derived live from demo kinematics, not recorded sensor data
First the corpus. Then the models built on it.
Capture and fidelity build one thing: a proprietary, action-labeled corpus of force-controlled assembly. Robot model architecture is commoditizing, open GR00T, open LeRobot, open π0. The data underneath it is not. The corpus is the moat; the specialist models are the compounding return on it.
The corpus cannot be scraped, only captured one interaction at a time, and it grows with every task we run. Each turn of the loop widens the data lead.
No model trains on Roborecs data yet, and no Roborecs model ships today. The model layer opens on two gates: the corpus crosses the scale published results show a specialist policy needs, and we add the ML research leadership to build it. We would rather state the gates than imply a model, or a team, that does not exist yet.
Our specialist models are task policies, not a robot, and they run on any LeRobot-compatible stack. The data is the product: any corpus a customer licenses is carved out of our own model roadmap, so we never ship a policy that competes with a data customer. The policies we do build prove what the corpus can produce.
EXAMPLE POLICIES · ROADMAP
Electronics assembly
A specialist vision-language-action policy for connector insertion, cable routing, and screwdriving. Post-trained on the force-torque and tactile channels of the assembly corpus, where a millimetre and a few newtons decide success. Delivered LeRobot-compatible.
Precision kitting
Bin-to-fixture part placement under tight tolerance. The same corpus, a different task family.
The models sit downstream of the open foundation models, not across from them. Post-trained on the corpus, delivered in the same LeRobot-compatible format. Compatible with, never competing with, NVIDIA Isaac GR00T N1.7 and Physical Intelligence π0.
NEXT
Run a pilot.
Tell us your robot and target tasks. We scope a capture program, agree a sample spec, and deliver an evaluation set before any volume commitment.