EN Submit a tool
Model

Black Forest Labs Releases FLUX 3 Multimodal Model, Supporting Single Generation of 20-Second Video with Native Audio

Published: Source: IT Home (RSS)

ShareXFacebookTelegramWhatsApp

Black Forest Labs has launched FLUX 3, a multimodal foundation model in Early Access, using a unified architecture to jointly learn images, videos, and audio. The model is built on the Self-Flow learning framework and can output up to 20-second videos with native audio in a single generation, supporting tasks such as text-to-video, image-to-video, and multi-shot concatenation.

Read the original (opens in a new tab)

News stream data aggregated by AI HOT

Related newsLatest in this category