volcengine / video-to-video-lip-sync

Faites parler vos vidéos dans votre langue, image par image — l'API de synchronisation labiale de Volcengine offre un lip sync vidéo à vidéo au pixel près et un doublage par IA en quelques minutes.

Pricing: 8 credits per second (~$0.04). High-tier top-ups (+10% bonus) bring effective pricing down to ~$0.036 per second.
Input

Click to upload or drag and drop

Supported formats: MP4, QUICKTIME, X-MATROSKA Maximum file size: 500MB

Video asset URL

Click to upload or drag and drop

Supported formats: MPEG, WAV, X-WAV, AAC, MP4, OGG Maximum file size: 10MB

Target pure vocal audio URL; used to drive video lip movements.

Service identifier

Enable vocal separation to suppress background noise.

Whether to enable scene segmentation and speaker identification. Supported only in Basic mode.

Supported in lite mode. Whether to loop the video when the audio is longer than the video.

Supported in lite mode. Whether to loop the video in reverse (backward). Requires align_audio to be set to true.

Supported in lite mode. Start time of the template video, in seconds.

Output
output typevideo

no output

README

Complete guide to using volcengine/video-to-video-lip-sync

API Video-to-Video Lip Sync de Volcengine : API de synchronisation labiale par IA et de doublage vidéo

Original image
Synchronisation labiale précise à l'image près
Pipeline de doublage vidéo multilingue
Intégration native à l'écosystème
Traitement asynchrone à haut débit des tâches

FAQ sur l'API Video-to-Video Lip Sync de Volcengine