Write the brief
Describe the output you need in plain words — subject, mood, format, aspect ratio. No prompt engineering, no camera crew.
MODEL FAMILY — LIPSYNC

HOW IT WORKS
Describe the output you need in plain words — subject, mood, format, aspect ratio. No prompt engineering, no camera crew.
Open the model picker and choose Lipsync. The studio shows the exact credit cost of the run before you commit.
Generate, review, change one line of the brief, and run it again — iteration costs a sentence, not a reshoot.
WHAT YOU GET
| No. | Capability | Detail |
|---|---|---|
| 01 | Talking-head presenter | Animate a portrait or brand avatar to deliver your voiceover with accurate lip-sync and blink. |
| 02 | Any language, any voice | Supply any language audio track — the model syncs lips to the waveform regardless of language. |
| 03 | Video lipsync | Reanimate an existing video clip with a new audio track — dub or revoice existing footage. |
| 04 | Fast turnaround | Most lipsync jobs complete in under 60 seconds — fast enough for same-day campaign deployment. |
REAL RENDERS
INCLUDED VARIANTS
| No. | Model | Capability |
|---|---|---|
| 01 | Lipsync | Lipsync |
The rate card
FAQ
Audio-driven lipsync: animate a portrait or video clip to match any voice track. On Orvik it sits in the Lipsync lane of the studio, on the same plan as every other frontier family.
Lipsync work, directed from a plain-language brief. The listed variants — Lipsync — all run from the same brief box in the studio, and every sample on this page is a real engine output.
Rates are per model and the studio shows the exact credit cost before you generate. The plan math — monthly credit allowances and what they convert to — is on /pricing.
Yes — 1 variant listed (Lipsync). The live studio reflects the current catalog, so new variants appear as the catalog updates.
Most lipsync models accept standard MP3 and WAV audio. Upload your voiceover or generated audio track, pair it with a portrait image, and generate a talking-head video.
Yes — generate a voiceover in the Text to Audio bucket first, then feed it into a lipsync model. The studio supports both steps end to end.
Lipsync models sync to the audio waveform regardless of language — so you can produce the same presenter speaking Hindi, English, or Tamil from separate audio tracks.
Keep exploring
LAST LOOKS
A closing reel of real engine output — video and image, across formats and use cases. Every tile is a sample render.