A peaceful pixelated lake valley holds every detail as the camera pushes in.
- Home
- Supported Models
- Flux 3
Updated August 2026
Flux 3 is here
Flux 3 generates video with native audio, up to 20 seconds in a single generation. Go text to video or image to video. Run it on Comfy Cloud today. Open weights coming soon.
Made with Flux 3
Three dancers, three rooms, one rhythm, timing locked across every panel.
Shattered porcelain reassembles on white, every fragment tracked in reverse.
A drone crosses from one environment to the next.
Outlaws ride into town, hooves and crowd noise with the frame.
A bear closes in at the checkpoint, game-engine look held steady.
Choose a plan
Access cloud-powered ComfyUI workflows with straightforward, usage-based pricing.
Start Comfy Cloud for free. Upgrade when you're ready.
5 free runs on real GPUs — no credit card required.
Billed monthly
- Included: 30 minute max workflow runtime
- Included: Add more credits anytime
- Not included: Import your own models
- Not included: Longer workflow runtime (up to 1 hr)
Generates ~380 5s videos*
Billed monthly
- Included: 30 minute max workflow runtime
- Included: Add more credits anytime
- Included: Import your own models
- Not included: Longer workflow runtime (up to 1 hr)
Generates ~670 5s videos*
Billed monthly
- Included: 30 minute max workflow runtime
- Included: Add more credits anytime
- Included: Import your own models
- Included: Longer workflow runtime (up to 1 hr)
Generates ~1,915 5s videos*
Built for teams collaborating on workflows together.
Billed monthly
Generates ~13,405 5s videos*
Everything in Pro, plus:
- Included: Invite members
- Included: Members can run workflows concurrently
- Included: Shared credit pool for all members
- Included: Role-based permissions
Coming soon...
- Coming soon: Shared workflows & assets
- Coming soon: Projects
Need more members? Looking for more flexibility or custom features?
*Based on 5s videos created with the Wan 2.2 Image-to-Video template using default settings (81 frames, 18fps, 640x640, 4-step sampler)
Q&A
No. Flux 3 interprets and rewrites your prompt before generation, so plain language works. Describe the scene the way you would brief a colleague. Keyword stacking and caption-style prose add nothing.
It rewrites, but the rewrite preserves what you explicitly specify. The more precisely you describe a scene, the more of the output is your call rather than the model's. Vague prompts hand control to the rewriter.
Quoted text without a visible speaker tends to render as on-screen text. To get a spoken line: quote the line, describe a speaker visible on camera, and add "no on-screen text, no subtitles".
Describe audio in layers and name each one separately: ambient sound, music, and speech. Each lands as its own layer. Lumping them into one phrase gives you less control over the mix.
Yes. Structure the prompt as "SHOT ONE ... HARD CUT. SHOT TWO ..." and it produces real cuts inside a single generation. For an uncut take, ask explicitly for "one continuous unbroken shot".
Consecutive shots that are too similar will not register as edits. Change scale, location, or colour between shots so the cut reads cleanly.
Yes. Specify it directly, such as "one continuous music bed across all three shots", alongside the shot structure.
Three examples, copy and paste ready. Plain brief: A cozy ramen shop on a rainy Tokyo night: steam rising from the broth, neon reflections in the window puddles, the cook working calmly. Rain patter and quiet kitchen sounds. Spoken line: A weather presenter on camera in front of a stylized storm map, speaking directly to the lens: "Storm season is here and this time, we're ready." Confident delivery, clean studio lighting. No on-screen text, no subtitles. Multi-shot: SHOT ONE: wide aerial of a desert highway at dawn, a single red car speeding through. HARD CUT. SHOT TWO: interior close-up, the driver's hands drumming the wheel to the radio. HARD CUT. SHOT THREE: from the roadside, the car shrinking into the heat haze. Warm engine hum under one continuous music bed across all three shots.
Twenty seconds with sound, in one generation.
One engine, every way to run it
Run Flux 3 in the browser today. Batch campaigns with the API, or bring it in-house.
Comfy MCP: now turn your agent into a creative technologist.
Your AI assistant can access the ecosystem, build workflows, and generate images, video, audio, or 3D.