Video-pretrained world models reach robotics
Praxis-1 claims mostly third-person video pretraining transfers across robot embodiments, with a reported 0.95 simulation-to-real correlation. Event-referential grasping and Agent Priors-guided Policy Learning add temporal grounding, active view selection, and compositional structural priors; independent real-robot validation remains decisive.
Sources (4)
Updated Oct 4, 2026