Hunyuan Video text-to-video generator creates high-quality 720p videos with customizable resolution, aspect ratio, and frame count. Features pro mode for enhanced quality.
Added Mar 11, 2025
Approx. Price
$0.400 per video
Model Type
text-to-video
Settings
Generation controls available for this model.
Output Format
Default Duration
N/A
Aspect Ratio
Default
16:9
Options (2)
16:9 (Landscape), 9:16 (Portrait)
The aspect ratio of the generated video
Number of Frames
Default
129
Options (2)
129 frames, 85 frames
The number of frames in the generated video
Pro Mode
Default
No
Enable higher quality generation with 55 steps (2x cost)
Resolution
Default
720p
Options (3)
480p, 580p, 720p
The resolution of the generated video
Safety Checker
Default
No
Enables or disables the safety filter
Benchmarks
Benchmarks
No benchmark data is available yet for this model.
Examples
Loading examples…
Related video models
Compare Hunyuan Video with similar models from the same provider or model family.
Hunyuan Video 1.5
hunyuan-video-15Hunyuan Video 1.5 generates 5 or 8 second clips from text or an input image. Supports 480p/720p and landscape or portrait runs routed automatically based on whether an image is attached.
LTX-2.5 Fast
lightricks/ltx-2.5/fastSpeed-optimized audiovisual generation from text, an image, or a 2-20 second audio clip. Creates synchronized video and audio in one pass, with output up to 4K and optional start/end-frame control.
LTX-2.5 Pro
lightricks/ltx-2.5/proHigh-fidelity audiovisual generation from text, an image, or a 2-20 second audio clip. Creates polished synchronized video and audio in one pass, with 720p/1080p output and optional start/end-frame control.
Wan 3.0 Image-to-Video
alibaba/wan-3.0/image-to-videoAnimate a first-frame image into a cinematic video with optional last-frame guidance, synchronized audio, deep-thinking controls, and 2–30 second output.
Wan 3.0 Reference-to-Video
alibaba/wan-3.0/reference-to-videoReference-guided video generation using images, videos, and audio for subject consistency, motion, timing, and scene continuity, with 2–30 second output.
Wan 3.0 Text-to-Video
alibaba/wan-3.0/text-to-videoCinematic text-to-video generation with synchronized audio, deep-thinking prompt interpretation, 2–30 second duration, and 480p, 720p, or 1080p output.