FLUX icon

FLUX 3 Video

3 Video

Generate HD video clips up to 20 seconds with native audio from text prompts, starting images, keyframes or short source videos. Create coherent multi-shot scenes with camera changes, dialogue, sound effects and ambient audio in a single generation. Iterate with Draft Mode before rendering full-quality output, reducing cost and wait time during creative direction reviews.

Reviewed by ToolWorthy Editors·updated today·3 Video released 15 days ago

Pricing:From $0.06/per second
Jump to section
FLUX 3 Video generation interface with prompt controls, keyframes, timeline and native audio lanes

More tools to compare

Neurona AI icon

Neurona AI

SenseNova U1 icon

SenseNova U1

Dreamina AI icon

Dreamina AI

Kittl Image Generator icon

Kittl Image Generator

Phot.AI Pattern Generator icon

Phot.AI Pattern Generator

Z-Image icon

Z-Image

Pros & Cons

Pros

  • Native audio-video generation — Frames, dialogue, effects and ambience are generated together, reducing separate sound-design steps for first drafts.
  • Flexible input modes — Text, images, keyframes and short video continuation support different creative workflows.
  • Longer clip length — Up to 20-second generations make the model useful for short ads, explainers and multi-shot concepts.
  • Multi-shot prompting — Scene and camera-angle changes can be created within one clip instead of assembled manually from isolated outputs.
  • Draft Mode — Faster previews help teams approve motion and composition before rendering final quality.

Cons

  • Paid per second — Video output can become expensive for long clips or repeated iteration.
  • Limited initial scope — Some FLUX 3 capabilities, including expanded references, FLUX 3 Image and FLUX 3 Dev, remain on the roadmap.
  • Vendor benchmarks only — Published performance claims are useful but should not replace independent testing.
  • Partner differences — Feature access may vary between the BFL API and third-party platforms.
  • Resolution-tier complexity — Draft, HD and FHD modes have different prices and limits, so teams should test the right tier before production use.

Overview

FLUX 3 Video is Black Forest Labs' first video-generation release from the FLUX 3 multimodal model family. Black Forest Labs announced FLUX 3 on July 23, 2026, and the current FLUX 3 model page makes video generation available through the BFL dashboard and API, moving FLUX beyond still-image generation into clips with motion, audio, dialogue and scene continuity.

Compared with FLUX.2 [klein], which focused on fast image generation and editing, FLUX 3 Video changes the product surface: users can now generate up to 20-second clips in HD or FHD, continue existing clips, define keyframes and produce native audio with the frames. This makes FLUX more relevant to teams comparing AI video generators for ads, storyboards, explainers and short-form production.

What's New

Video and Audio Generation Become Generally Available

Black Forest Labs introduced FLUX 3 on July 23, 2026 as a multimodal foundation model trained across images, video and audio. Its current product page exposes FLUX 3 Video through the BFL dashboard and API, with text-to-video, image-to-video, keyframe generation, continuation and native audio workflows.

The model creates clips up to 20 seconds long in HD or FHD resolution bands. Native audio is generated alongside the video, so dialogue, sound effects and ambience are part of the same generation path rather than a separate post-production pass.

Text, Images, Keyframes and Video Continuation

FLUX 3 Video supports several input modes for different creative starting points. Text-to-video handles simple or detailed prompts, image-to-video can animate a starting frame, and keyframe-to-video connects defined moments into a controlled transition.

The release also adds video continuation. Users can provide an existing clip, then tell FLUX 3 what should happen next. The model uses the input context to continue movement, camera behavior, dialogue and sound.

Multi-Shot Scene Construction

FLUX 3 Video can create multiple scenes and camera angles within a single generation while keeping the sequence coherent. That matters for short ads, explainers and storyboards where a single static shot is not enough.

This is a practical step beyond previous image-focused FLUX releases. Instead of generating separate stills and stitching them manually, teams can prompt for a short sequence with camera changes, motion, audio and dialogue in one pass.

Native Dialogue, Audio and Multilingual Output

The release includes dialogue, sound effects and ambient audio generation. Black Forest Labs lists support for English dialects plus Chinese, Spanish, French, German, Japanese, Portuguese, Russian, Italian, Indonesian, Turkish, Hindi, Punjabi and more, with lip-syncing.

For creators, this reduces the number of tools needed to produce a first cut. A prompt can define the scene, movement, spoken line, mood and language, then FLUX 3 Video generates both the frames and the sound layer together.

Draft Mode for Faster Iteration

Draft Mode lets users explore creative directions before paying the time and cost of a full-quality generation. A draft returns a faster preview at lower cost; when the draft is approved, FLUX 3 renders the final version with the same subjects, composition and motion.

This is especially useful for teams that review motion concepts with clients or internal stakeholders. Instead of committing to full-resolution outputs for every idea, they can narrow down timing, shot logic and composition first.

Performance Benchmarks

Black Forest Labs published preliminary human-preference results with the July 23 announcement. In those early evaluations, FLUX 3 was preferred over Grok Imagine Video in up to 69% of comparisons, Kling v3 Pro in 60%, Happy Horse v1 in 59%, Happy Horse 1.1 in 57%, Seedance 2.0 and Gemini Omni Flash in 52%, Runway Gen-4.5 in 77%, and Luma Ray 3.2 in 93%.

BFL explicitly described those results as early and expected further improvements during rollout. Treat the numbers as vendor-published directional evidence, not independent benchmarks, and validate the model against a team's own prompts, languages, styles and motion requirements.

Availability & Access

FLUX 3 Video is available through the BFL API and selected partners. The model page also lists ecosystem partners including Canva, Cloudflare, fal, Krea, Magnific, OpenArt, OpenRouter, Picsart, Runway, Envato and HeyGen, indicating that access may depend on the platform where a team already works.

The broader FLUX 3 family is still rolling out. Black Forest Labs says future releases will expand video controllability, add generation from combinations of image, video and audio references, and ship FLUX 3 Image plus FLUX 3 Dev as an open-weight variant. Teams that need image generation, local deployment or open weights should treat those as roadmap items rather than assume they are included in this Video release.

Current Limitations

  • FLUX 3 Video is only one part of the FLUX 3 rollout; FLUX 3 Image, FLUX 3 Action and FLUX 3 Dev are separate release tracks.
  • API and partner availability may differ by platform, account type and rollout stage.
  • Draft Mode is HD only; standard generations support HD and FHD pricing tiers.
  • Vendor benchmarks should be validated against a team's own prompts, languages and motion needs.

Pricing & Plans

Black Forest Labs uses pay-as-you-go pricing for FLUX generation. The FLUX 3 model page lists text or image to video at $0.06/sec for Draft HD, $0.17/sec for standard HD, and $0.29/sec for standard FHD. Video-to-video costs more: $0.12/sec for Draft HD, $0.41/sec for standard HD, and $0.53/sec for standard FHD. A 5-second standard text-to-video clip in HD is shown as $0.85 before any platform-specific fees or enterprise discounts.

For high-throughput workloads, Black Forest Labs offers enterprise pricing with volume discounts, SLA guarantees and dedicated support. The same pricing page also lists open-weight licensing plans for FLUX.2 models, but those plans do not imply open-weight access to FLUX 3 Video; BFL describes FLUX 3 Dev as a future release.

Best For

  • Creative teams producing short ads, social clips or campaign concepts that need motion and audio in the same draft.
  • Agencies building client storyboards where multi-shot continuity matters more than a single cinematic frame.
  • Product marketers creating explainer videos, app mockups or branded motion-design tests.
  • Educators and publishers making short documentary or tutorial clips from compact prompts.
  • Developers comparing API-based video generation models for text-to-video, image-to-video and continuation workflows.

FAQ

Is FLUX 3 Video the same as FLUX 3 Image?

No. FLUX 3 is the broader multimodal model family, while FLUX 3 Video is the video-generation release from that family. Black Forest Labs says FLUX 3 Image and FLUX 3 Dev are separate parts of the rollout.

How long can FLUX 3 Video clips be?

FLUX 3 Video can generate clips up to 20 seconds long. The public model page lists HD at up to 1 megapixel per frame and FHD at up to 2 megapixels per frame; Draft Mode is HD only.

Does FLUX 3 Video generate audio?

Yes. FLUX 3 Video generates native audio with the frames, including dialogue, sound effects and ambient sound. It also supports multilingual dialogue with lip-syncing across a broad set of languages.

Can FLUX 3 Video continue an existing video?

Yes. Users can provide an existing video clip, then prompt what should happen next. The model uses the input context to continue movement, camera behavior, dialogue and audio.

Is FLUX 3 Video open weight?

No public open-weight FLUX 3 Video release is listed for this milestone. Black Forest Labs describes FLUX 3 Dev as a future open-weight variant for content creation and action prediction.

Version History

3 Video

Current Version

Released on July 23, 2026

+What's new
3 updates
  • Generate HD video clips up to 20 seconds with native audio from text prompts, starting images, keyframes or short source videos.
  • Create coherent multi-shot scenes with camera changes, dialogue, sound effects and ambient audio in a single generation.
  • Iterate with Draft Mode before rendering full-quality output, reducing cost and wait time during creative direction reviews.

.2 [klein]

Released on January 15, 2026

View Update
+What's new
3 updates
  • Introduce 4B and 9B parameter models with 4-step distillation for sub-second inference, making interactive image generation practical on supported cloud and local hardware
  • Deliver a unified architecture for text-to-image generation, single-reference editing and multi-reference composition without switching between separate FLUX model families
  • Release 4B variants under Apache 2.0 with open weights, enabling commercial self-hosting paths for teams that need local deployment and licensing flexibility

.2

Released on November 25, 2025

View Update
+What's new
3 updates
  • Combine up to 10 reference images in a single generation to maintain consistent characters, products, styles and scenes across complex creative projects like e-commerce campaigns and brand materials
  • Generate images up to 4MP resolution with significantly improved typography, perfect for high-quality advertisements, posters and detailed product photography
  • Control quality and speed with flexible step parameters in FLUX.2 [flex], allowing you to quickly draft concepts at low steps then refine winners at high steps for optimal cost efficiency

.1 Kontext

Released on May 29, 2025

+What's new
3 updates
  • Prompt with both text and images to extract and modify visual concepts from existing photos or generated images, enabling rapid variations like changing product backgrounds or updating clothing while preserving subjects
  • Edit images iteratively with multi-turn instructions up to 8x faster than competing models, supporting progressive refinement workflows for portrait retouching and complex scene building
  • Test FLUX capabilities without technical integration using the new BFL Playground, perfect for POCs, team evaluations and product demos before API implementation

.1 [pro]

Released on October 2, 2024

+What's new
3 updates
  • Generate images 6x faster than FLUX.1 [pro] while improving image quality, prompt adherence and output diversity, enabling commercial-scale batch production for advertising and content platforms
  • Access transparent per-image pricing through the generally available BFL API at 4 cents per image for FLUX1.1 [pro], simplifying cost forecasting for enterprise budgets
  • Benefit from 2x speedup on updated FLUX.1 [pro] without changing outputs, allowing existing workflows to scale immediately

.1

Released on August 1, 2024

+What's new
3 updates
  • Choose between three variants tailored for different needs - [pro] for maximum quality, [dev] for downloadable non-commercial weights, and [schnell] for ultra-fast local iteration under Apache 2.0 license
  • Generate images across flexible aspect ratios and resolutions from 0.1 to 2.0 megapixels, accommodating everything from thumbnails to high-definition marketing materials
  • Leverage 12B parameter architecture built on flow matching with rotary positional embeddings and parallel attention layers, providing the technical foundation for speed and stability improvements

Top alternatives

Related categories

From the blog

View all →

Track FLUX in ToolWorthy Weekly

Important tool updates, better alternatives, and selected AI signals in one weekly brief.

Weekly only. Unsubscribe anytime.