logo
RunComfy
  • ComfyUI
  • TrainerNew
  • Models
  • API
  • Pricing
discord logo
MODELS
Explore
All Models
LIBRARY
Generations
MODEL APIS
API Docs
API Keys
ACCOUNT
Usage

FLUX 3 Video: Text-to-Video with Native Audio on Models and API | RunComfy

blackforestlabs/flux-3/text-to-video

FLUX 3 Video creates text-to-video clips up to 20 seconds with optional native audio at 720p or 1080p.

The text prompt describing the video you want to generate.
Aspect ratio of the generated video. auto lets the model choose.
Resolution of the generated video.
Duration of the generated video in seconds. auto lets the model choose.
Whether to generate audio for the video.
The safety tolerance level for the generated video. 0 is the strictest and 4 is the most permissive.
Idle
The rate is $0.17 per second at 720p, and $0.29 per second at 1080p.

Introduction To FLUX 3 Video

Black Forest Labs' FLUX 3 Video turns a written prompt into a cinematic clip with optional synchronized native audio at 720p or 1080p, for up to 20 seconds per generation.
Trading timeline edits, separate dubbing, and frame-by-frame cleanup for prompt-driven motion and sound, FLUX 3 Video helps marketers, agencies, and previz artists ship short clips faster.
For developers, FLUX 3 Video on RunComfy can be used both in the browser and via an HTTP API, so you don't need to host or scale the model yourself.
Ideal for: High-Conversion Video Ads | Shot-Accurate Film Previsualization | Multi-language Lip-Synced Brand Narratives

Black Forest Labs / FLUX 3 Video#


FLUX 3 Video is Black Forest Labs' text-to-video model that turns a written brief into a short cinematic clip, with optional synchronized native audio in the same pass. On RunComfy you can iterate in the model UI and reuse the same settings over HTTP API.


Output format: Video / up to 20 seconds / 720p or 1080p / optional native audio


Highlights#


  • Prompt to picture: Describe the shot, action, and mood; FLUX 3 Video invents the motion without a timeline edit.
  • Sound in the same pass: Enable audio to get speech, ambience, and effects timed to the picture, or turn it off for silent footage.
  • Length room: Choose auto duration or lock 5–20 seconds so dialogue beats and slower dramatic shots have space.
  • Delivery framing: Aspect ratios from ultrawide 21:9 to vertical 9:16, plus auto when you want the model to pick.
  • Two resolutions: Generate at 720p by default, or step up to 1080p when delivery needs sharper detail.
  • Safety dial: Adjust safety tolerance from strict (0) to more permissive (4) for the scene you need.

Parameters#


ParameterRequiredTypeDefaultRange / OptionsDescription
prompt*Yes (*)string——Text describing the video you want to generate.
aspect_ratioNostringautoauto, 21:9, 2:1, 16:9, 4:3, 1:1, 3:4, 9:16Output frame shape; auto lets the model choose.
resolutionNostring720p720p, 1080pOutput resolution.
durationNostringautoauto, 5–20Clip length in seconds; auto lets the model choose.
generate_audioNobooleantruetrue, falseWhether to generate synchronized audio.
safety_toleranceNointeger20–40 is strictest; 4 is most permissive.

Pricing#


Billing is based on the length of the generated clip and the selected resolution.


ResolutionRate
720p$0.17 per second
1080p$0.29 per second

Estimated cost examples (720p)


DurationApprox. cost
5 s~$0.85
10 s~$1.70
20 s~$3.40

Estimated cost examples (1080p)


DurationApprox. cost
5 s~$1.45
10 s~$2.90
20 s~$5.80

How to Use#


1) Open FLUX 3 Video on RunComfy and write a clear prompt for subject, camera, motion, and sound.

2) Set Aspect Ratio to match delivery (for example 16:9 landscape or 9:16 vertical), or leave auto.

3) Choose Resolution: start at 720p while iterating, then switch to 1080p for final delivery.

4) Pick Duration as a fixed length (5–20 s) or leave auto so the model decides.

5) Keep Generate Audio on when you want speech, SFX, or ambience; turn it off for silent video.

6) Optionally adjust Safety Tolerance if the scene needs a stricter or looser filter.

7) Generate, review the clip, then refine the prompt or duration before scaling up.

8) For API use on RunComfy, send the same parameters; results appear in your job history.


Prompt & Reference Tips#


  • Name camera and motion explicitly (medium close-up, slow push-in, handheld vs locked-off).
  • With audio on, mention dialogue tone or ambient sound so the soundtrack matches the shot.
  • Prefer one continuous action over contradictory beats in a single prompt.
  • Start with shorter fixed durations while iterating, then extend once the take looks right.
  • Use auto aspect ratio when framing is flexible; lock a ratio for known social or film placements.
  • Keep prompts concrete: subject, setting, lighting, and pace beat vague mood words alone.
  • If motion feels noisy, simplify the prompt and regenerate at the same settings.

How FLUX 3 Video compares to other models#


  • Compared with image-only FLUX editions, FLUX 3 Video adds temporal motion and optional native audio under one foundation.
  • Compared with multimodal video tools that also accept image or video references, this endpoint focuses on text-to-video with audio control, aspect ratio, and resolution.
  • Based on publicly available information, FLUX 3 Video is positioned for expressive faces, physically grounded motion, and dialogue-friendly clips with sound locked to on-screen events.
  • Ideal use case: short ads, film previz, and branded narratives where a strong written brief is enough to drive the shot.

More Models to Try#


  • Seedance 2.0 for multimodal text, image, video, and audio reference workflows
  • Kling video models for alternative motion styles and aspect-ratio options
  • Veo text-to-video when you want another native-audio generation path
  • FLUX image models when you need stills before moving into video

Official Resources#


  • Black Forest Labs

In short, FLUX 3 Video on RunComfy turns prompts into cinematic clips with optional native audio at 720p or 1080p, ready for browser iteration and API production.

Related Models

wan-2-2/animate/video-to-video

Transforms input clips into synced animated characters with precise motion replication.

gemini-omni-flash/video-edit

Edit a source video from a text instruction while keeping scene coherence.

veo-3-1/reference-to-video

Create rapid high-quality video drafts with precise style and speed

flux-3/text-to-video/draft

FLUX 3 Draft: Fast, low-cost text-to-video previews at 720p

happyhorse-1.0/image-to-video

HappyHorse 1.0 I2V on Alibaba animates a still image into native 1080p video with physics-accurate motion and identity-stable subjects.

dreamina-3-0/pro/text-to-video

Turn text into detailed cinematic scenes with Dreamina 3.0 precision.

Frequently Asked Questions

What is FLUX 3 Video used for?

FLUX 3 Video turns a text prompt into a short cinematic clip with optional synchronized native audio. It targets ad creative, film previsualization, and branded storytelling where prompt-driven motion and sound matter.

How long can a FLUX 3 Video clip be?

You can set duration to auto or lock a fixed length from 5 to 20 seconds. Auto lets the model choose; fixed lengths give you predictable pacing for dialogue and dramatic beats.

Does FLUX 3 Video support native audio?

Yes. Generate Audio is on by default, so the model can output synchronized speech, SFX, and ambience with the picture. Turn it off when you need silent video.

What resolution and aspect ratio options does FLUX 3 Video support?

Resolution options are 720p (default) and 1080p. Aspect ratio defaults to auto, or you can lock 21:9, 2:1, 16:9, 4:3, 1:1, 3:4, or 9:16 for known delivery formats.

What input fields does FLUX 3 Video require?

Only the prompt is required. Optional fields are aspect_ratio, resolution, duration, generate_audio, and safety_tolerance (0–4, default 2).

How does safety_tolerance work in FLUX 3 Video?

Safety tolerance controls how strict content filtering is. 0 is the strictest setting and 4 is the most permissive; the default is 2.

Will FLUX 3 Video offer open weights or LoRA fine-tuning?

Black Forest Labs has announced FLUX 3 editions including hosted video, image, action prediction, and a planned FLUX 3 Dev open-weight backbone. Once that backbone lands, community LoRA workflows for recurring characters and brand looks are the natural next step.

How do I move from testing FLUX 3 Video in the browser to production API integration?

Prototype in the RunComfy model UI, then call the same template through the RunComfy API using identical Input fields (prompt, aspect_ratio, resolution, duration, generate_audio, safety_tolerance). Validate prompts in the UI first, then use your account API key and credits for automated jobs.

How much does FLUX 3 Video cost on RunComfy?

Billing is based on generated video duration and resolution: $0.17 per second at 720p, and $0.29 per second at 1080p. See the pricing shown on this page for the latest rates.

Follow us
  • LinkedIn
  • Facebook
  • Instagram
  • Twitter
Support
  • Discord
  • Email
  • System Status
  • Affiliate
Video Models
  • MiniMax H3 Open
  • FLUX 3 Image to Video
  • MiniMax H3 Open Image to Video
  • Wan 2.6 Flash
  • Happy Horse 1.1 reference to video
  • Seedance 1.5 Pro Text to Video
  • View All Models →
Image Models
  • Seedream 5.0 Pro Image Edit
  • Flux 2 Flash Edit
  • Nano Banana Pro
  • seedream 4.0
  • GPT Image 2
  • Qwen Image Edit 2511 LoRA
  • View All Models →
Legal
  • Terms of Service
  • Privacy Policy
  • Cookie Policy
RunComfy
Copyright 2026 RunComfy. All Rights Reserved.

RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.

Examples Of FLUX 3 Video

Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...