Image-and-audio avatar video generation for speech-driven talking-head clips, with 720p and 1080p output.
Added May 12, 2026
Approx. Price
$0.125 per video
Model Type
image-to-video
Settings
Generation controls available for this model.
Output Format
Default Duration
N/A
Resolution
Default
720p
Options (2)
720p, 1080p
Output video resolution
Seed
Default
-1
Control reproducibility (-1 for random)
Video Prompt
Default
The person is talking.
Controls body movement, framing, and atmosphere
Benchmarks
Benchmarks
No benchmark data is available yet for this model.
Examples
Loading examples…
Related video models
Compare P-Video Avatar with similar models from the same provider or model family.
P-Video Animate
pruna-ai/p-video/animateMotion-control video generation that animates a reference image using movement from a source video, with optional prompt guidance, audio preservation, frame-rate control, and 720p or 1080p output.
P-Video Image-to-Video
pruna-ai/p-video/image-to-videoFast image-to-video generation with 1-20 second durations, 720p and 1080p output, and optional audio.
P-Video
pruna-ai/p-video/text-to-videoFast text-to-video generation with 1-20 second durations, 720p and 1080p output, optional audio, and common aspect ratios.
LongCat Avatar 1.5
wavespeed-ai/longcat-avatar-1.5Upgraded audio-driven talking or singing avatar generation from a single image with sharper lip sync and faster generation. Supports 480p/720p output up to 30 seconds.
LongCat Avatar 1.5 Multi
wavespeed-ai/longcat-avatar-1.5/multiAudio-driven two-person avatar generation from a single image and left/right audio tracks. Supports simultaneous or sequential dialogue, 480p/720p output, and up to 30 seconds of audio.
LongCat Avatar
longcat-avatarAudio-driven talking or singing avatar generation from a single image with lip-synced motion and consistent identity. Supports 480p/720p output up to 2 minutes.