DreamZero is a World Action Model that learns robot control by jointly predicting future world states and actions using video diffusion. It achieves over 2× better generalization to new tasks and environments compared to Vision-Language-Action models, and can adapt to new robot embodiments with just 30 minutes of play data while maintaining zero-shot generalization capabilities.