Speed-optimized audiovisual generation from text, an image, or a 2-20 second audio clip. Creates synchronized video and audio in one pass, with output up to 4K and optional start/end-frame control.
Added Aug 11, 2026
Approx. Price
$0.200 per video
Model Type
both
Settings
Generation controls available for this model.
Output Format
Default Duration
6
8 duration options
Aspect Ratio
Default
auto
Options (3)
Auto, Landscape (16:9), Portrait (9:16)
Auto follows an input image; text-only generation defaults to 16:9.
Camera Motion
Default
N/A
Options (9)
Automatic, Dolly In, Dolly Out, Dolly Left +5 more
Optional camera movement.
Duration
Default
6
Options (8)
6 seconds, 8 seconds, 10 seconds, 12 seconds +4 more
Clip length for text-to-video and image-to-video. Audio-to-video follows the input audio length.
Frames Per Second
Default
25
Options (4)
24 FPS, 25 FPS, 48 FPS, 50 FPS
Output frame rate for text-to-video and image-to-video.
Generate Audio
Default
Yes
Generate synchronized audio with text-to-video or image-to-video.
Guidance Scale
Default
5
Prompt adherence for audio-to-video (1-50).
Resolution
Default
1080p
Options (4)
720p, 1080p, 1440p, 4K
Output video resolution. Audio-to-video uses 1080p.
Benchmarks
Benchmarks
No benchmark data is available yet for this model.
Examples
Loading examples…
Related video models
Compare LTX-2.5 Fast with similar models from the same provider or model family.
LTX-2.5 Pro
lightricks/ltx-2.5/proHigh-fidelity audiovisual generation from text, an image, or a 2-20 second audio clip. Creates polished synchronized video and audio in one pass, with 720p/1080p output and optional start/end-frame control.
LTX-2.3 Quality
ltx-2.3-qualityLTX-2.3 Quality routes text, image, audio, reference video, extend-video, video-to-HDR, and optional LoRA inputs to FAL. Supports provider output sizes such as landscape, portrait, square, and auto with native synchronized audio.
FLUX.3
flux-3Generate up to 20-second videos with native audio from a prompt, a start image, start/end frames, multiple keyframes, or a source clip. FLUX.3 chooses the matching workflow automatically from what you attach.
Pixelcut Video Background Remover
pixelcut/video-background-removalRemove video backgrounds with frame-by-frame AI segmentation and temporally consistent edges. Supports transparent output, preset solid backgrounds, custom RGB backgrounds, and common video formats.
Bernini R Video
bernini-r-videoBernini R video generation and editing. A single NanoGPT model id routes prompt-only requests to text-to-video, up to 5 image inputs to reference-to-video, video inputs to edit-video, and image plus video inputs to reference-edit-video.
Luma Ray 3.2
luma/agent/ray/v3.2Cinematic text-to-video and image-to-video generation with strong motion control, optional reference images, seamless loops, 540p/720p/1080p output, and 5 or 10 second clips.