Skip to main content

AI lip-sync for existing faces

Lock mouth motion to dialogue with models typed lipsync in config. Use it on plates you already shot or generated.

Kling LipSync · Prompt: Diner window portrait, dialogue-ready framing

What is AI lip sync?

Lip-sync models align mouth shapes on a face video (or still-driven clip) to a dialogue track. FairStack lists type lipsync only on this page.

Lip Sync models

2 models matched from config for this capability.

MuseTalk 1.5

$0.0015/sec

MuseTalk 1.5 is a lip synchronization model that adds natural mouth movement to existing images or video, billed per second of output. The model specializes in lip sync only, driving mouth movements from audio input without generating full body motion or head movement, keeping the processing focused. Per-second billing makes it practical for high-volume production, batch processing, and applications where hundreds or thousands of clips need lip synchronization. The model works with both static images and existing video. MuseTalk 1.5 is the lip-sync-only option: premium lip sync models like Sync Lipsync 2.0 Pro and full talking head models like Kling Avatar add facial and body animation and are priced accordingly. Best suited for lip sync at scale, adding speech to portrait photos, and high-volume video lip synchronization where cost per second matters most. Available on FairStack at infrastructure cost plus a 20% platform fee; the current per-second rate is shown on this page.

How to use lip sync

Step 1

Provide face media

Video plate or a model that accepts image + audio.

Step 2

Add dialogue audio

Dry voice works better than mixed beds.

Step 3

Run a lipsync model

Picker is type=lipsync from config.

Step 4

Review phoneme edges

Re-run with cleaner audio if consonants smear.

Frequently asked questions

Which models appear? +

Config type lipsync — including Kling LipSync and MuseTalk 1.5 among the selectable set.

Still have questions? We're here to help.