FLUX 3 Video by Black Forest Labs supports text-to-video, image-to-video with 1–10 ordered keyframes, video continuation, and draft enhancement. Generate synchronized audio with 5–20 second HD or FHD output, configurable safety tolerance, and results stored on the APIMart CDN.
Create a video from a prompt
Choose an integer from 5 to 20 seconds.
Generate synchronized native audio with the video.
Generate a lower-cost, low-quality HD draft first, then convert it to the final video with a second request.
Stored on the APIMart CDN
Transparent pricing with no hidden fees. Pay only for what you use.
* Actual costs are subject to final output.
Generate from a prompt, animate 1–10 ordered keyframes, continue a source video, or enhance a completed draft. Choose 5–20 seconds, HD or FHD output, and synchronized native audio.
50K+
Active Users
99.9%
Uptime
2x
Faster
70%
Cost Savings
Confirmed generation modes, output controls, and delivery features
Practical workflows built from the four confirmed generation modes
Submit and retrieve a FLUX 3 Video task in three steps
Select text-to-video, ordered image keyframes, video continuation, or enhancement of a completed draft task.
Add the required prompt or source input, then choose 5–20 seconds, HD or FHD, aspect ratio, audio, and safety tolerance.
Submit the asynchronous task, poll its status, and retrieve the completed video from the APIMart CDN.
Common questions about FLUX 3 Video capabilities and workflows
FLUX 3 Video is Black Forest Labs' video model for text-to-video, ordered image keyframes, video continuation, and draft enhancement with optional synchronized native audio.
It supports text-to-video, image-to-video with ordered keyframes, video continuation, and enhancement of a completed draft task.
Yes. Synchronized native audio is enabled by default. Set audio to false when you need a silent video; disabling audio does not change the price.
Requests accept an integer duration from 5 to 20 seconds. Output resolution can be HD or FHD, while draft mode is limited to HD.
Use t2v for a prompt, i2v for ordered images, v2v for a source video, and draft_enhance with a completed draft task ID.
You can submit 1–10 images in order. One image is the start frame; with two or more, the first and last are the endpoints and intermediate images are distributed between them.
Create an HD draft with draft set to true. After it completes, submit its task ID with draft_from_task_id to render a final HD or FHD video.
Set safety_tolerance from 0 to 4. When a task completes, the result URL is transferred to the APIMart CDN for long-term access.
Explore more models in the same category.

Gemini Omni Flash Preview
gemini-omni-flash-preview is a multimodal video generation and editing model launched by Google.

Kling 3.0 Turbo
Kling-3.0-Turbo: A high-speed, high-quality AI video generation model ideal for quickly creating short videos.

Pixverse V6
pixverse-v6 is PixVerse's sixth-generation AI video generation model, primarily used for text-to-video and image-to-video generation.

Omni Flash Ext
Omni-Flash-Ext is an extended video generation model in version 4.6.4 of Google's Gemini series.