Tagged “#world-models”
-
Human Motion as a Control Interface for Interactive World Models
How 2D pose, SMPL bodies, and tracked head-and-hand motion become causal guidance for responsive video world models rather than offline animation controls.
-
From Human Video to Robot Action: Five Interfaces Across the Embodiment Gap
Five interfaces that turn human video into robot supervision: latent actions, interaction tokens, digital twins, contact topology, and robotized demonstrations.
-
Streaming Autoregressive Video Generation
A research review of streaming autoregressive video generation: causal distillation, self-forced rollouts, visual memory, few-step sampling, and world models.
-
Embodied AI World Models
A practical map of embodied world models covering prediction targets, action grounding, robot data, uncertainty, planning, and closed-loop deployment.
-
On 3D, Video World Models, and the Approaching ImageNet Moment of Perception-Action Learning
How 3D structure and video world models could support an ImageNet-scale shift from isolated vision tasks toward causal perception-action learning.
See all tags.