The Rundown AI homepage

Independent tool overview

FLUX 3 at a glance

FLUX 3 is Black Forest Labs' multimodal model family. Its first public product, FLUX 3 Video, generates or continues 5–20 second video clips with synchronized speech, sound effects, and ambience from text, images, keyframes, or an existing clip. The video endpoint is live, but the documentation still calls it a preview and the broader image, action, editing, and open-weight roadmap is not fully released.

Visit the official FLUX 3 site ↗
FLUX 3 product preview
Available product
FLUX 3 Video API and playground
Input modes
Text, 1–10 images or keyframes, and video continuation
Output
5–20 second video with optional synchronized audio
Resolution
HD or upscaled Full HD at 24 fps
Pricing unit
Per second of output
Model status
Generally accessible endpoint documented as preview
Reviewed
August 31, 2026

Overview

What FLUX 3 is

FLUX 3 is Black Forest Labs' attempt to use one underlying architecture across image, video, audio, language, and eventually action prediction. The publicly usable FLUX 3 Video endpoint currently turns text into video, animates one to ten pinned images or keyframes, and continues an existing video while generating synchronized audio alongside the frames.

The live API supports clips at 24 frames per second in HD or upscaled Full HD, with common landscape, square, portrait, and ultrawide aspect ratios. Text-to-video and image-to-video run for 5–20 seconds; video continuation runs for 5–15 seconds. Audio is on by default, and the documentation highlights multilingual speech, lip sync, effects, ambience, in-scene typography, multiple scenes, and varied visual styles.

Draft mode is unusually practical for cost control. It creates a cheaper HD preview and returns a cache bundle that can be enhanced into the same approved shot at full quality, instead of paying for a new high-quality interpretation that may change composition or motion. Result links expire after roughly two hours, so applications need reliable polling, download, storage, and failure handling.

The broader FLUX 3 announcement is ahead of the currently shipped product. Image synthesis and editing, richer video editing and Omni Reference, FLUX 3 Action, private weights, and an open-weight FLUX 3 Dev model remain staged releases, selected-partner access, or roadmap items. Do not describe the entire multimodal research model as generally available merely because the video-generation endpoint is live.

Use cases

Who FLUX 3 is best for

The strongest fit depends on the job you need the product to complete, not the size of its feature list.

Short-form creative video

Studios and creators generating five- to twenty-second concept shots, social clips, motion designs, stylized scenes, or ad previsualization with native sound.

Storyboard and keyframe control

Teams that want to pin opening, closing, or intermediate images to exact moments rather than relying only on a text prompt.

Developer integrations

Products that can manage asynchronous jobs, expiring result URLs, content moderation, rights checks, per-second cost controls, and human review.

Capabilities

Core FLUX 3 features

1

Text-to-video

Creates a new clip from a written scene description, including synchronized speech, effects, ambience, and optional multi-shot direction.

2

Image-to-video and keyframes

Uses one to ten images as pinned frames or storyboard references, with optional timestamps to define when particular frames must occur.

3

Video continuation

Accepts an existing clip and generates the next 5–15 seconds while attempting to carry forward momentum, framing, scene logic, and sound.

4

Native audio

Generates multilingual speech, lip sync, environmental sound, effects, and ambience jointly with the video; audio can be disabled when a separate sound workflow is preferred.

5

Draft and enhance

Produces a lower-cost HD preview with a reusable draft cache, then upgrades the selected draft without reinterpreting the approved shot.

6

API and playground

Supports browser experimentation and an asynchronous API that returns a job identifier, polling URL, signed output link, and short retrieval window.

Process

How the FLUX 3 workflow works

  1. Step 1

    Clear every source asset

    Confirm rights and consent for prompts, faces, voices, performances, music, footage, reference images, brands, and locations before uploading anything.

  2. Step 2

    Choose the narrowest input mode

    Use text for exploration, one or more keyframes for visual control, and continuation only when rights and continuity from an existing clip matter.

  3. Step 3

    Iterate in Draft

    Test prompts at the cheaper draft rate, keep seeds and parameters, compare motion and audio defects, and enhance only the approved cache bundle.

  4. Step 4

    Download immediately

    Poll asynchronously with backoff, retrieve completed files before the roughly two-hour signed URL expires, store them in controlled infrastructure, and handle failed or timed-out jobs.

  5. Step 5

    Run a human release review

    Check faces, hands, physics, dialogue, lip sync, text, trademarks, provenance, disclosure, consent, audio rights, accessibility, and misleading-realism risk before publication.

Cost

FLUX 3 pricing and free plan

BFL uses credits where one credit equals $0.01, but FLUX 3 Video is easiest to understand as per-second pricing. Text-to-video and image-to-video share one rate; continuation costs more. Drafts are HD only. Playground and API pricing are the same, and batches or repeated attempts multiply the bill.

Text-to-video or image-to-video draft

$0.06/second

Fast HD previews for 5–20 second generations.

  • About $0.30 for 5 seconds
  • About $1.20 for 20 seconds
  • Returns a draft cache for deterministic full-quality enhancement
  • Fresh drafts remain separate paid generations

Text-to-video or image-to-video HD

$0.17/second

Full HD-tier render at the documented hd resolution and 24 fps.

  • About $0.85 for 5 seconds
  • About $3.40 for 20 seconds
  • Native audio included unless disabled
  • One to ten images supported for image-to-video

Text-to-video or image-to-video FHD

$0.29/second

Full render finished at the documented Full HD dimensions through the video upsampler.

  • About $1.45 for 5 seconds
  • About $5.80 for 20 seconds
  • FHD is an upscaled output path
  • Optional standalone 2K and 4K upscaling costs extra

Video continuation draft

$0.12/second

Cheaper HD exploration for a 5–15 second continuation of an existing clip.

  • About $0.60 for 5 seconds
  • About $1.80 for 15 seconds
  • Source-video upload and rights required
  • Audio and scene logic continue with the generated segment

Video continuation full

$0.43/second HD or $0.54/second FHD

Production-quality continuation for a 5–15 second output.

  • HD costs about $2.15–$6.45 across the allowed duration
  • FHD costs about $2.70–$8.10
  • Storage, retries, moderation, and downstream processing are separate
  • Enterprise agreements and private access can use different commercial terms

Pricing checked . Check current pricing at the source ↗

Assessment

FLUX 3 strengths and limitations

Where it stands out

  • Generates video and synchronized audio together instead of requiring separate dialogue, effects, ambience, and lip-sync passes.
  • Text, keyframes, multiple pinned images, and continuation modes offer more control than a text-only clip generator.
  • Up to 20 seconds in one request is useful for short-form sequences and reduces the number of joins needed for a finished social clip.
  • Draft-to-enhance preserves the approved preview while reducing the cost of prompt exploration.
  • One asynchronous endpoint and clear per-second pricing make the live video capability straightforward to integrate and budget.

What to consider

  • The documentation still calls FLUX 3 a preview model even though the initial video endpoint is broadly available; behavior, prices, limits, and output quality can change quickly.
  • The announced FLUX 3 Image, broader editing, Omni Reference, action prediction, private weights, and open-weight Dev capabilities are not all generally available today.
  • FHD is produced through an upsampler rather than native Full HD generation, and 2K or 4K delivery requires a separate paid upscaling step.
  • Result URLs expire after roughly two hours. An application that does not download promptly can lose access and may need to pay for regeneration.
  • Generated people, dialogue, text, physics, brands, music, and factual scenes can be wrong or misleading. The developer terms require human evaluation and prohibit unauthorized deepfakes and impersonation.
  • Owning the output as between the user and BFL does not clear third-party copyright, music, performance, trademark, privacy, publicity, or biometric rights.
  • BFL's August 2026 non-EU self-serve API terms grant a perpetual license to use inputs and outputs for operation, product development, and model training. Playground terms offer a prospective training opt-out by email outside beta programs, while negotiated and EU terms may differ.

Compare

FLUX 3 alternatives

The right alternative depends on the specific output, workflow, controls and budget your project requires.

Content Creator

Runway Gen-4.5

Choose Runway Gen-4.5 for a broader creator-facing video production environment and a mature suite of generation, editing, and control tools.

Explore Runway Gen-4.5

Content Creator

Kling 3.0

Choose Kling 3.0 to compare native-audio generation, longer-form consistency, motion, and consumer-app workflows.

Explore Kling 3.0

Content Creator

Veo 3.1

Choose Veo 3.1 when Google ecosystem access, Flow tooling, native audio, or Vertex AI integration is a better operational fit.

Explore Veo 3.1

Questions

FLUX 3 FAQs

Is FLUX 3 available now?

FLUX 3 Video is available through the BFL playground, API, and selected partners. The live endpoint supports text-to-video, image or keyframe-to-video, and video continuation. Other announced FLUX 3 modalities and weights remain preview, selected-access, or roadmap items.

How long can FLUX 3 videos be?

Text-to-video and image-to-video support 5–20 second outputs. Video continuation supports 5–15 seconds. All current modes output 24 fps, with HD or upscaled FHD full renders; drafts are HD.

Does FLUX 3 generate audio?

Yes. Audio is enabled by default and can include multilingual speech with lip sync, effects, and ambience generated alongside the frames. Developers can disable audio, but must separately clear any rights in uploaded or generated music, voices, recordings, and performances.

How much does a 20-second FLUX 3 video cost?

For text-to-video or image-to-video, a 20-second draft is about $1.20, a full HD render about $3.40, and a full FHD render about $5.80. Retries, batches, storage, continuation, and optional upscaling add cost.

What is FLUX 3 Draft mode?

Draft mode creates a faster, cheaper HD preview and returns a draft-cache bundle. Sending that bundle to draft_enhance produces the same selected shot at full quality, unlike a new direct render that can reinterpret the prompt.

Can I use FLUX 3 output commercially?

BFL's API guidance includes commercial output rights, and its developer terms say the user owns output as between the parties. That is not rights clearance: users still need permission for source footage, music, performers, likenesses, trademarks, confidential material, and the intended distribution.

Does BFL train on FLUX 3 API inputs and outputs?

The August 4, 2026 non-EU self-serve API terms grant BFL rights to use inputs and outputs to improve and train its systems. The general playground terms describe a prospective email opt-out outside beta programs, and EU or negotiated enterprise terms can differ. Review the contract that governs the exact endpoint before sending confidential assets.

Bottom line

Our FLUX 3 verdict

FLUX 3 Video is a compelling short-form API for teams that value native audio, pinned keyframes, continuation, and a cost-efficient draft-to-enhance loop. The page should sell what is live—not the entire research roadmap. For production, budget repeated generations, download outputs immediately, verify every frame and sound, clear source and likeness rights, and resolve BFL's input/output training terms before using confidential media.

Visit FLUX 3 website ↗
The Rundown University

AI training for the future of work.

Get access to all our AI courses, hundreds of real-world AI use cases, live expert-led workshops, an exclusive network of AI early adopters, and more.

AI Courses

Get unlimited access to all of our current & upcoming industry-specific AI courses for the duration of your subscription.

Daily Guides

To keep up with the rapid pace of AI, our team publishes AI implementation guides daily. Our library contains 300+ practical use cases to automate real-world work.

Workshops

Join weekly, live, interactive sessions with industry leaders who are at the forefront of AI for hands-on implementation guidance and exclusive insights.

Community

Network with an exclusive community of AI-first professionals who are working smarter with AI. Learn how early adopters are using AI in their work and businesses.