Kling Lip Sync on Kanvora
Synchronize mouth movement in one private 2-10 second 720p or 1080p source video to one private 2-60 second audio file through fal. No text prompt is sent. fal documents $0.014 per input video second rounded up to five-second increments, so Kanvora's existing credit formula charges 6 credits for a 2-5 second video or 12 credits for a 6-10 second video.
- Output
- 720p/1080p source-framed
- Credits
- 6-12 credits
- Input
- 2-10s video + 2-60s audio
- Video
- 720p or 1080p · MP4/MOV · up to 100 MB
- Audio
- MP3/WAV · up to 5 MB
- Billing
- $0.014/input video second, rounded to 5s
- Expected wait
- 1-8 minutes
- Provider path
- fal Kling lip sync
- Registry verified
- Aug 22, 2026
Start with the exact private source contract, not a blank upload.
This is a direct handoff to Kanvora's real creator, not an embedded generator. The model and operation are already attached, and no text prompt will be sent. Upload both private files after opening the creator.
Open the generatorNo text prompt · private source video + private audio
Official Kanvora examples
Site-owned AI-generated proofs from this model, never a public user gallery. Open an example to carry its cataloged direction into the creator.
From model page to stored proof
One public decision path, followed by the real authenticated workflow.
- 1
Prepare a short source video
Use a 2-10 second MP4 or MOV up to 100 MB at 720p or 1080p. Keep one face clear enough to inspect mouth movement after the request.
- 2
Prepare the replacement audio
Use a 2-60 second MP3 or WAV up to 5 MB. Kanvora verifies its format and duration privately before submission.
- 3
Review the video-duration billing bucket
The audio length does not set this endpoint's price. A verified 2-5 second input video costs 6 credits; a 6-10 second input video costs 12 credits.
- 4
Inspect timing and identity before download
Review lip timing, teeth, jaw motion, facial identity, audio alignment, and frame stability. Lip sync is not a guarantee of perfect phoneme timing or identity preservation.
Compare this live operation
Choose by the live Kanvora path, not by a model-family claim from somewhere else.
| Model | Resolution | Credits | Provider path | Use it when |
|---|---|---|---|---|
| Kling 2.6 Motion Control | Source-framed | 18-175 credits | fal Kling motion control | Transfers body and camera motion from one private video to a character image and requires a text direction. |
| Kling Lip Sync | 720p/1080p source-framed | 6-12 credits | fal Kling lip sync | Replaces mouth timing from private audio without sending a text prompt; 2-10 second source video only. |
| Veo 3.1 Extend | 720p | 267-267 credits | Google direct Veo video | Continues only an eligible owned Veo output rather than synchronizing supplied audio. |
Frequently asked questions
Current Kanvora behavior, stated without extending the model's capability boundary.
How many credits does Kling Lip Sync use?
fal documents $0.014 per input video second and rounds the billed duration up to the next five-second increment. Kanvora applies its existing $0.012 supplier-cost coverage per credit, so a 2-5 second video costs 6 credits and a 6-10 second video costs 12 credits.
What files can I upload?
Kanvora accepts one private MP4 or MOV source video from 2 through 10 seconds, at 720p or 1080p, up to 100 MB. It also accepts one private MP3 or WAV audio file from 2 through 60 seconds up to 5 MB.
Does the audio duration change the credit cost?
No. The documented supplier price is based on input video seconds rounded to five-second increments. Kanvora still verifies the audio's 2-60 second duration before submission.
Does Kling Lip Sync need a prompt or guarantee perfect sync?
No prompt is sent to this audio-to-video endpoint. The endpoint is designed to synchronize mouth movement, but every output still needs review for lip timing, identity drift, facial artifacts, and audio alignment.
Take this model into the real creator.
The prepared prompt and legal model parameters go with you. Sign-in remains at the generation boundary.