Generative video models have made rapid progress in the past two years, moving from short, noisy clips to near-photorealistic scenes. This article reviews the main methods behind modern AI video generation, focusing on diffusion-based approaches and temporal modeling. We also discuss current limits in motion, physics, and long-form generation, and outline directions for future work.