Video with native audio
Black Forest Labs says FLUX 3 Video can generate video and audio together, so dialogue, ambience, and physical events can be created in the same pass.
Explore what FLUX 3 Video offers, then start creating with our advanced AI models below.
Announced July 23, 2026 · Early Access
FLUX 3 is Black Forest Labs' multimodal foundation model for image, video, audio, and action prediction. FLUX 3 Video is its audiovisual generation path: one model trained to connect how scenes look, move, and sound instead of treating each medium as an isolated tool.
This site is an independent FLUX 3 AI Video guide and creator. The available generator above uses the third-party model shown in its selector; it does not present current output as official FLUX 3 output.
Official Early Access preview
These are announced capabilities, not a claim that every control is already available in the independent FLUX 3 Video Generator on this page.
Black Forest Labs says FLUX 3 Video can generate video and audio together, so dialogue, ambience, and physical events can be created in the same pass.
The announced workflows include text-to-video, image-to-video, video-to-video, video and audio continuation, and keyframe-controlled transitions.
Early Access materials describe diverse video with native audio up to 20 seconds in one generation, with multi-shot chaining for longer sequences.
FLUX 3 Video is designed for multilingual dialogue, typography, animation, cinematic footage, candid camera styles, and multiple aspect ratios.
Available now on this platform
You can use the creator above today for prompt-driven and reference-driven video. The model selector names the actual available workflow, currently Wan 2.5 Fast or Wan 2.5 Quality.
The official video model remains in Early Access. When a supported FLUX 3 integration becomes available, it can be added as its own clearly labeled model rather than silently rebranding another provider's output.
Simple workflow
Write the subject, action, environment, camera movement, lighting, pacing, and sound you want.
Use the independent creator above with its clearly labeled Wan 2.5 Fast or Quality workflow and the input mode that fits your source.
Submit the job, wait for processing, review the selected model's result, and download the video when it is ready.
Creative applications
Explore framing, movement, atmosphere, and scene rhythm before a production shoot.
Turn a product image or written concept into a short motion study for campaign planning.
Test landscape, square, and vertical creative directions before committing to a final edit.
Use references and prompts to explore character motion, transitions, and stylized sequences.
Release map
BFL's staged plan covers video and audio APIs and private weights, image synthesis and editing, selected action-prediction partnerships, and an open-weight multimodal backbone called FLUX 3 Dev. FLUX 3 Dev Video searches reflect interest in that future open-weight path, not a generally available download on this site.
FLUX 3 Video is the video and audio generation capability of Black Forest Labs' FLUX 3 multimodal foundation model. BFL announced text-to-video, image-to-video, video-to-video, continuation, keyframe, multilingual dialogue, and native-audio workflows.
FLUX 3 Video entered Early Access on July 23, 2026. Black Forest Labs describes a staged rollout of video and audio APIs and private weights after Early Access, so general availability and access may change.
No. Flux 3 Video is an independent information and creation platform. It is not affiliated with, endorsed by, or operated by Black Forest Labs.
No. The current creator uses the third-party model named in its model selector. We do not label those videos as FLUX 3 output. A FLUX 3 option should only be added after a real supported integration is available.
The current tool supports text-guided and reference-driven video workflows, including a start image, first and last frames, multiple reference images, or a text prompt. Available controls depend on the selected model.
FLUX 3 Dev is the announced open-weight access path for the multimodal backbone. BFL says it is planned for content creation across video, audio, and image as well as action prediction, but release details are still evolving.
FLUX.2 focuses on image generation and editing. FLUX 3 expands the foundation into jointly trained image, video, audio, and action capabilities, with FLUX 3 AI Video serving as its audiovisual creation path.
Create with an available model
Use text, a starting frame, multiple references, or a first-and-last-frame setup. Every generated result remains labeled by the actual model used.
Open the AI video creator