Floyo
Floyo
Workflows
API
Pricing
Floyo
Floyo
Workflows
API
Pricing

Sync Lipsync 2 Pro · Video to Video

Upload a video and a new audio track, and Sync Lipsync 2 Pro replaces the lip movements to match the new audio while preserving facial detail, skin texture, and identity.

41

Generates in about 5 mins 34 secs

Nodes & Models

SyncLipsync_floyo
VideoToFrames
LoadAudio
LoadVideo
FloyoStickyNote
VHS_VideoCombine

ABOUT THE WORKFLOW

Sync Lips to New Audio
Upload a video of someone talking and an audio file of the new dialogue. The model replaces the mouth movements in the video to match the new audio, frame by frame. The rest of the face, body, and scene stay untouched. Works on live-action footage, 3D animation, and AI-generated video up to 4K.

Partner node. This workflow calls an external API, so each run uses credits from your API wallet. No API key needed. Floyo handles the connection.

Model

  • Lipsync 2 Pro by Sync Labs. A video editing model with diffusion-based super resolution that generates lip sync while preserving fine facial detail like teeth, beards, and freckles, with zero speaker-specific training required.


HOW IT WORKS

Step 1. Upload your video
A video of someone talking. The speaker must be actively moving their mouth in the input. Static or still faces do not produce lip movement.
Works great with: live-action footage · AI-generated talking videos · 3D animation · dubbed content

Step 2. Upload your audio
The new dialogue, narration, or vocal track you want the speaker to lip-sync to. Any language works.
Works great with: voiceovers · translated dialogue · re-recorded lines · singing tracks

Step 3. Hit run and download
Lipsync 2 Pro replaces the mouth movements in the video to match the new audio and returns an MP4 with the synced result.
Ready for: Premiere · DaVinci Resolve · After Effects · TikTok · YouTube · broadcast

First time? Upload a video of someone talking and your new audio. Leave every setting as-is. The defaults (lipsync-2-pro, bounce mode, temperature 0.5) handle most cases.


RECOMMENDED SETTINGS

Quick-start guide. Find the goal that matches yours and copy the settings.

  • Standard dialogue replacement — lipsync-2-pro, bounce sync mode, temperature 0.5, no emotion override. No changes needed.

  • More expressive mouth movement — Raise temperature toward 0.8. The model produces wider, more dynamic lip shapes. Good for energetic dialogue or singing.

  • Subtler, calmer lip movement — Lower temperature toward 0.3. The model produces tighter, more restrained mouth movement. Good for narration or quiet speech.

  • Multi-person scene with one speaker — Turn on active_speaker_auto_detect. The model identifies who is talking and syncs only that person's lips.

  • Speaker's mouth is partially hidden — Turn on occlusion_detection_enabled. The model handles obstructions like microphones, hands, or hair near the mouth. Note: this slows processing.

  • Lip sync not landing on the audio — Make sure the input video shows the speaker actively talking. Lipsync 2 Pro learns the speaking style from the input and needs visible mouth movement to work. Static faces produce no lip motion.

Prompt: No prompt is needed. The workflow is driven by the video and audio inputs. Upload both and run.


LEARN

📹 Videos

✨ Quick links


USE CASES

🌐 Video Dubbing and Localization
Replace the original dialogue with a translated audio track and get lip-synced output in the new language without reshooting.

🎬 Dialogue Replacement in Post-Production
Swap a line reading, fix a flubbed take, or replace placeholder audio with final voice acting while keeping the original performance intact.

📱 Social Media Content Repurposing
Re-voice an existing video with new dialogue, a different language, or a different speaker for a new audience or platform.

🎮 Character Voice Replacement
Swap the voice on animated or AI-generated characters without re-rendering. Works on 3D animation, anime, and AI-generated talking videos.


WHAT WORKS BEST / WHAT TO AVOID

✅ Works great

  • Videos where the speaker is actively talking with visible mouth movement

  • Clean audio with isolated speech and minimal background noise

  • Front-facing or slight-angle shots where the mouth and jaw are visible

  • Live-action, 3D animation, and AI-generated talking videos

⚠️ May produce softer results

  • Still frames or video segments where the speaker is not moving their mouth

  • Extreme profile views where the lips are barely visible

  • Very fast or heavily accented speech with rapid phoneme changes

  • Videos where hands, microphones, or hair cover the lower face (unless occlusion detection is on)


FAQ

What is Sync Lipsync 2 Pro?
Lipsync 2 Pro is a video editing model by Sync Labs that replaces lip movements in an existing video to match a new audio track. It uses diffusion-based super resolution to preserve fine facial detail like teeth, beards, and freckles during the edit. It works on live-action footage, 3D animation, and AI-generated video at up to 4K resolution, with no speaker-specific training required.

Does Sync Lipsync 2 Pro work with any language?
Yes. The model adapts to any language automatically. It reads acoustic features from the audio waveform and generates matching mouth movements without language-specific training. This makes it suitable for multilingual dubbing and localization workflows.

What does the temperature setting control?
Temperature controls how expressive the lip movements are. At 0.3, the model produces subtle, restrained mouth movement suited for calm narration. At 0.8, it produces wider, more dynamic lip shapes suited for energetic speech or singing. The default of 0.5 works for most content.

Can I use Sync Lipsync on a video where the person is not talking?
Lipsync 2 Pro requires the input video to show active speaking motion. The model learns the speaker's style from their existing mouth movement and replaces it. Static or still faces do not produce lip sync. If you are working with AI-generated video, include "person is speaking naturally" in the generation prompt to ensure the character's lips are moving.

What is the difference between Lipsync 2 Pro and Sync 3?
Lipsync 2 Pro uses diffusion-based super resolution for fine detail preservation and generates faces at 512x512, which works well for most 1080p video. Sync 3 generates at native 4K resolution with built-in obstruction detection and can open silent lips to match audio. Sync 3 costs more per second but handles extreme angles and occlusion better.

Is Sync Lipsync 2 Pro licensed for commercial use?
Lipsync 2 Pro is a proprietary model by Sync Labs. Commercial use is governed by their terms of service. Review the current terms for your specific use case, especially for broadcast, advertising, and client-facing content.

How to run AI lip sync for video online?
You can run AI lip sync for video online through Floyo. No installation, no setup, no API key to wire up. Open the workflow in your browser, upload your video and audio, and hit run. Free to try.


WHY FLOYO?

Floyo is the only platform with team collaboration for ComfyUI in the browser. You run workflows with no install. You share run history, assets, and models across your team. You pay only when you generate. Floyo supports open-source and closed-source models.

A designer runs an edit and likes the result. A teammate opens that exact run from shared history and keeps going. No file handoffs. No version confusion.

For studios and enterprise teams, Floyo adds private workspaces, pooled resources, and a team usage dashboard. Other ComfyUI cloud tools run for one person at a time. Floyo runs for the whole team, with transparent per-generation costs.


Ready to try it?
Upload a video of someone talking, upload the new audio, and hit run.

→ Launch Workflow, Free

Questions? Watch the free course or check the FAQ above.

Read more

N