Switch language한국어
Back to the list

FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching

TL;DR AI

Key summary

2 min read
  1. Researchers introduced FlowLong, an inference-only method for much longer video generation without extra training.

  2. It stitches overlapping sliding-window predictions with Tweedie matching to preserve temporal consistency and visual stability.

  3. A stochastic early-sampling phase followed by deterministic sampling helps reduce drift and improve quality over long sequences.

  4. FlowLong is architecture-agnostic and can extend the native length of different video diffusion models, including audio-video and text-to-3DGS settings.

Read the original