Tagged “#robotics”
-
From Human Video to Robot Action: Five Interfaces Across the Embodiment Gap
Five interfaces that turn human video into robot supervision: latent actions, interaction tokens, digital twins, contact topology, and robotized demonstrations.
-
Embodied AI World Models
A practical map of embodied world models covering prediction targets, action grounding, robot data, uncertainty, planning, and closed-loop deployment.
-
From π0 to π0.7: A Tutorial on Open-pi and Robot Foundation Models
How Physical Intelligence's open-pi models turn language and vision into continuous robot actions, then add context, memory, feedback, and steering.
-
On 3D, Video World Models, and the Approaching ImageNet Moment of Perception-Action Learning
How 3D structure and video world models could support an ImageNet-scale shift from isolated vision tasks toward causal perception-action learning.
See all tags.