MODEL //
ONE PLANNER. ONE EXECUTOR.
Two layers with different jobs: a deliberate system that comprehends multi-step work,
and a video-pretrained system that turns each step into motion.
fig. 01 · the two-system architecture
The data
01 Pretraining: 150,000 hours of video spanning 2,000+ real-world tasks.
02 Grounding: first-person human demonstrations before any robot data.
03 RL on our own hardware, aimed at the failure modes deployments expose.
Where it works, domain by domain: use cases →
Technical report on training is coming soon.