Switch language한국어
Back to the list

Black Forest Labs Releases FLUX 3: A Multimodal Flow Model for Image, Video, Audio and Robot Action Prediction

TL;DR AI

Key summary

2 min read
  1. Black Forest Labs launched FLUX 3, a multimodal foundation model built on Self-Flow.

  2. The model is trained jointly on images, video, and audio, and includes early-access video and action capabilities.

  3. A related policy model, FLUX-mimic, suggests the same backbone could also support robot action prediction.

  4. The release highlights a single shared architecture for generative media and robotics, with potential gains in cross-modal learning.

Read the original