Browse all Mirelo Ai video models

Mirelo SFX1.6 Video to Audio

mirelo-ai/sfx1.6/video-to-video

Mirelo SFX1.6 Video to Audio

mirelo-ai/sfx1.6/video-to-video

Generate synchronized sound effects for an uploaded video and return the clip with a new audio track. Supports source videos up to 60 seconds.

Added May 22, 2026

Starting Price

From $0.050 per video

Final price depends on the selected settings and is shown before generation.

Model Type

video-to-video

Settings

Generation controls available for this model.

Output Format

N/A

Default Duration

N/A

Seed

Number

Default

-1

Use -1 for random output or set a seed for repeatable variations.

Benchmarks

No benchmark data is available yet for this model.

Examples

Loading examples…

Compare Mirelo SFX1.6 Video to Audio with similar models from the same provider or model family.

MiniMax H3 Max Lip Sync

minimax/h3-max/lip-sync/image-to-video

Animate a portrait or character image from supplied audio with transcription-guided lip sync. Audio must be at least 5 seconds; longer inputs are clipped to the first 15 seconds. Supports talking and singing clips from 480p through 2K.

Wan 3.0 Prime Video Edit

alibaba/wan-3.0-prime/video-edit

Accelerated Wan 3.0 editing for an existing video with optional image or audio references. Uses the first 15 seconds of the source and supports 2–15 second output at 480p, 720p, or 1080p.

Wan 3.0 Prime Video Extend

alibaba/wan-3.0-prime/video-extend

Accelerated video extension that appends 2–30 seconds with optional target last-frame guidance. Source audio is preserved, and clips longer than 120 seconds use their final 120 seconds as context.

Wan 3.0 Video Edit

alibaba/wan-3.0/video-edit

Edit an existing video with natural-language instructions and optional image or audio references. Uses the first 15 seconds of the source and supports 2–15 second output at 480p, 720p, or 1080p.

Wan 3.0 Video Extend

alibaba/wan-3.0/video-extend

Extend an existing video by 2–30 seconds with optional target last-frame guidance. Source audio is preserved, and clips longer than 120 seconds use their final 120 seconds as context.

Seedance 2.5 Talking Avatar

bytedance/seedance-2.5/talking-avatar

Turn a portrait and voice recording into an expressive talking-avatar video with natural lip sync, facial motion, and optional direction for gestures or presentation style. Supports up to 120 seconds at 480p or 720p.