Flux 3 AI prompts that generate video with native audio. Copy split screen, dialogue, and up to 20 second clip prompts directly.
2026.07.28
Flux 3 is the multimodal model Black Forest Labs announced on July 23, 2026. It learns from images, video, and audio inside a single unified architecture, so a text prompt alone can produce a clip that already has sound. Video (FLUX 3 Video) is open first in early access, with image generation to follow. Carat's prompt gallery collects AI video prompts that have been tested against real output.
The clearest difference in Flux 3 is that video and audio are generated together instead of being stitched afterward. A single generation can run up to 20 seconds, and dialogue and sound effects come with the result. It supports text-to-video, continuing from a starting frame, carrying a character or element from a reference clip into a new scene, and filling the gap between keyframes. This makes it useful for cinematic video and dialogue scenes.
The examples the community shared most right after launch split the frame into two or four panels and show the same event from different camera angles at once. People, object positions, lighting, and timing stay aligned across the panels. Fixed cameras and camera movement can be mixed in one frame without the action drifting apart, which suits CCTV style footage or incident reconstructions that need several viewpoints at the same time.
You can describe a situation without writing the lines, and the model will generate the conversation and burn subtitles onto the frame. It handles dialogue in multiple languages, and voices and appearances tend to hold across a 20 second clip. That is worth knowing for lip sync video or interview style content where speaking is the point.
Flux 3 responds better to prompts that break a scene into time blocks with camera placement than to a single short description. Every prompt in the gallery below is registered together with its actual output, so you can copy it directly or swap the subject. To compare results against another video model, look at the Seedance 2.5 prompts as well.