WM3D: Native-3D World Modeling for Robot Action
Published:
WM3D argues that robot world models should predict explicit 3D state — depth, point geometry, pose, robot state, task text, and action-conditioned dynamics — then render video as a view of that imagined world.
