FAL Lip Sync 2.0

FAL Lip Sync 2.0 is the middle tier of FluxoKit's lip sync ladder, resyncing a face video to a connected audio track with more tolerance for quick speech than 1.9. It runs in the browser from 30 credits for a 5-second clip, billed by duration, on plans starting at $14 per month with no fal.ai account.

Examples

6 generations

Real generations with FAL Lip Sync 2.0

Like what you see? Make the next one yourself. Plans from $14/month.

Create with FAL Lip Sync 2.0

Overview

FAL Lip Sync 2.0 sits between the budget 1.9 and the premium 3.0 rungs. It takes the same two inputs, a source face video and the audio to match, and holds sync through faster dialogue and more movement than the entry tier, which is where 1.9 starts to drift.

FluxoKit runs it in the browser with credit pricing and no fal.ai account. A 5-second clip costs from 30 credits and scales with length at the provider's per-minute rate, so 2.0 is the tier you reach for when 1.9 is not quite holding but you do not yet need the top rung's price.

Create with FAL Lip Sync 2.0Plans from $14/month. Your first paid subscription includes a 30-day satisfaction guarantee, subject to the Refund Policy.

AI lip sync for pace and movement

This is the working default for real dubbing and revoicing: interviews, ad reads, and creator content where the speech has pace and the head moves, but the shot is still broadly front-facing. Connect the original video and the replacement audio and 2.0 rewrites the mouth to match, holding the sync where the budget tier would slip.

Chain it after a text-to-speech node to drive a generated voiceover onto a face, or feed it a translated audio track to revoice existing footage. Keep 1.9 for the simplest heads and escalate to 3.0 only for the punishing cases, so you pay the middle rate for the middle problem.

Where FAL Lip Sync 2.0 earns its credits

Ad and content teams use 2.0 as the everyday sync tier: the quality is high enough for published talking-head video, and at from 30 credits for a 5-second clip it stays far cheaper than the top rung when the footage does not demand it. Every run shows its exact credit cost in your history, so a large dubbing batch stays predictable.

FAL Lip Sync 2.0 vs 3.0, 1.9, and ElevenLabs Video Dubbing

FAL Lip Sync 3.0

The top rung wins on the hardest sync: fast speech, side profiles, heavy motion. Step up from 2.0 only when a shot actually needs it, since 3.0 costs more per clip.

FAL Lip Sync 1.9

The budget rung is cheaper per clip but drifts on quick speech. Use 1.9 for simple heads and 2.0 once the dialogue picks up pace.

ElevenLabs Video Dubbing

Dubbing translates and revoices a full video. Lip Sync 2.0 only aligns the mouth to audio you supply; chain dubbing into a sync pass when you need both.

Capabilities

Video support

Yes

Audio support

Yes

Source video input

Yes

Source audio input

Yes

Lip sync

Yes

Audio-driven lip sync

Yes

Output type

Video

Default duration

5 s

Sync modes

cut_off, loop, bounce, silence, remap

Default sync mode

Cut_off

Parameters and inputs

Each field below shows what the model accepts and the limits to apply.

4 parameters

Parameters

Video

video

File
Required
Yes
Default
Not set
Accepts
video/*

Audio

audio

File
Required
Yes
Default
Not set
Accepts
audio/*

Sync Mode

sync_mode

Select
Required
No
Default
cut_off

Options

  • Cut Off (cut_off)
  • Loop (loop)
  • Bounce (bounce)
  • Silence (silence)
  • Remap (remap)

Video Duration Seconds

video_duration_seconds

Number
Required
No
Default
5
Min
0.1
Step
0.1

Frequently asked questions

›How much does FAL Lip Sync 2.0 cost per clip?

It costs from 30 credits for a 5-second clip on FluxoKit and scales with clip length at the provider's per-minute rate. Plans start at $14 per month and every run shows its exact credit cost in your history.

›How is FAL Lip Sync 2.0 different from 1.9?

2.0 holds sync through faster speech and more head movement than 1.9, at a higher per-clip rate. Use 1.9 for simple front-facing heads and 2.0 once the footage gets harder.

›What inputs does FAL Lip Sync 2.0 need?

A source video with a visible face and an audio track to match. Connect both in the FluxoKit canvas and the model rewrites the mouth movement to the audio.

›Do I need a fal.ai account to use FAL Lip Sync 2.0?

No. FluxoKit runs it and more than 100 other AI models under one subscription with credit based pricing and a 30-day money-back guarantee.

›How are FAL Lip Sync 2.0 examples documented?

The inputs, sync modes, and per-run credits on this page come from FluxoKit's live model registry and pricing. Worked before-and-after clips with transcripts arrive with the video showcase wave; nothing here claims examples that are not yet published.

Start creating

Run FAL Lip Sync 2.0 in your browser. No API key, no provider account.

The everyday lip sync tier for real talking-head video. One studio, more than 100 AI models.

  • More than 14,000 generations delivered
  • Your first paid subscription includes a 30-day satisfaction guarantee, subject to the Refund Policy.
  • Checkout through Stripe, cancel anytime

Plans from

$14/month

Credit based pricing, live USD checkout

Create with FAL Lip Sync 2.0Or keep browsing the model docs