Extend Veo 3.1 videos by 7 seconds per call with smooth motion, preserved style, and strong scene coherence. Input must be Veo 3.1 generated. Supports up to 20 extensions for max 148 seconds total. 16:9 or 9:16 aspect ratio, 720p or 1080p.
Added Dec 15, 2025
Approx. Price
$2.80 per video
Model Type
video-to-video
Settings
Generation controls available for this model.
Output Format
Default Duration
N/A
Resolution
Default
1080p
Options (2)
720p, 1080p
Output resolution (must match input)
Benchmarks
Benchmarks
No benchmark data is available yet for this model.
Examples
Loading examples…
Related video models
Compare Veo 3.1 Extend with similar models from the same provider or model family.
Veo 3.1 Fast Extend
veo3-1-fast-extendFast video extension for Veo 3.1 clips. Adds 7 seconds per call with optimized speed for quick iteration. Input must be Veo 3.1 generated. Supports up to 20 extensions for max 148 seconds total. 16:9 or 9:16, 720p or 1080p.
Veo 3.1 Fast
veo3-1-fast-videoFast Veo 3.1 generation for text-to-video, image-to-video, and reference-to-video with up to 3 total images. Supports optional end frame control for image-to-video, native audio, and 4/6/8 seconds at 720p or 1080p.
Veo 3.1 Lite
veo3-1-lite-videoUnified Veo 3.1 Lite entry for text-to-video, image-to-video, and first/last-frame video generation. Supports 4/6/8 seconds at 720p or 1080p.
Veo 3.1
veo3-1-videoText-to-video and image-to-video with optional end frame control. Native audio generation, cinematic realism, and consistent subjects. Supports 4/6/8 seconds at 720p or 1080p.
Veo 3 Fast
veo3-fast-videoGoogle's fast Veo 3 model. Creates high-quality 8-second videos from text or images. Supports audio generation ($1.60 with audio, $1.20 without). Supports 16:9 and 9:16 aspect ratios. For best results, prompts should be descriptive and clear.
Veo 3
veo3-videoGoogle's latest Veo 3 model. Creates high-quality 8-second videos from text or images. Supports audio generation ($4.80 with audio, $3.20 without). For best results, prompts should be descriptive and clear. Include the subject, context, action, style, camera motion, composition, and ambiance details. Note: This model has strict content filters and may reject NSFW or sensitive content — we issue refunds for content policy rejections.