Switch language한국어
Back to the list

Wonder: Video World Model Done Better

TL;DR AI

Key summary

2 min read
  1. Researchers introduced Wonder, a real-time video world model that turns an image or conditional video into a playable scene.

  2. Users can navigate scenes with camera motion, re-render video-conditioned content, and preserve coherent appearance, geometry, and motion.

  3. The system uses dense coordinate-field camera conditioning, sparse attention memory, and improved distillation to better follow controls and keep diversity.

  4. Wonder can generate minute-scale videos at 16 FPS, advancing interactive simulation and long-horizon video creation.

Read the original