Bernini R video generation and editing. A single NanoGPT model id routes prompt-only requests to text-to-video, up to 5 image inputs to reference-to-video, video inputs to edit-video, and image plus video inputs to reference-edit-video.
Added Jun 9, 2026
Approx. Price
$0.240 per video
Model Type
both
Settings
Generation controls available for this model.
Output Format
N/A
Default Duration
5
3 duration options
Duration
Default
5
Options (3)
3 seconds, 5 seconds, 8 seconds
Estimated output length for billing and UI estimates.
Benchmarks
Benchmarks
No benchmark data is available yet for this model.
Examples
Loading examples…
Related video models
Compare Bernini R Video with similar models from the same provider or model family.
LTX-2.5 Fast
lightricks/ltx-2.5/fastSpeed-optimized audiovisual generation from text, an image, or a 2-20 second audio clip. Creates synchronized video and audio in one pass, with output up to 4K and optional start/end-frame control.
LTX-2.5 Pro
lightricks/ltx-2.5/proHigh-fidelity audiovisual generation from text, an image, or a 2-20 second audio clip. Creates polished synchronized video and audio in one pass, with 720p/1080p output and optional start/end-frame control.
FLUX.3
flux-3Generate up to 20-second videos with native audio from a prompt, a start image, start/end frames, multiple keyframes, or a source clip. FLUX.3 chooses the matching workflow automatically from what you attach.
Pixelcut Video Background Remover
pixelcut/video-background-removalRemove video backgrounds with frame-by-frame AI segmentation and temporally consistent edges. Supports transparent output, preset solid backgrounds, custom RGB backgrounds, and common video formats.
Luma Ray 3.2
luma/agent/ray/v3.2Cinematic text-to-video and image-to-video generation with strong motion control, optional reference images, seamless loops, 540p/720p/1080p output, and 5 or 10 second clips.
LTX-2.3 Quality
ltx-2.3-qualityLTX-2.3 Quality routes text, image, audio, reference video, extend-video, video-to-HDR, and optional LoRA inputs to FAL. Supports provider output sizes such as landscape, portrait, square, and auto with native synchronized audio.