Model
Black Forest Labs Releases FLUX 3 Multimodal Model, Supporting Single Generation of 20-Second Video with Native Audio
Black Forest Labs has launched FLUX 3, a multimodal foundation model in Early Access, using a unified architecture to jointly learn images, videos, and audio. The model is built on the Self-Flow learning framework and can output up to 20-second videos with native audio in a single generation, supporting tasks such as text-to-video, image-to-video, and multi-shot concatenation.
Read the original (opens in a new tab)
News stream data aggregated by AI HOT