How it works

Four moves from raw audio to a global release.

No tool-juggling, no manual handoffs. Every step picks up exactly where the last one left off — and every stage inherits the timecodes from the one before it, so nothing drifts between languages.

Localization pipeline 4 stages 80+ languages
Stage 01
Upload

Drop your video — any format, any source.

Camera footage, podcast feed, or edit-suite export: everything works. The file is the only thing the pipeline needs to start.

If you already have a transcript or a translation, upload it alongside your media. The pipeline skips the stage you've already done and the credit cost drops automatically, so you only pay for the work that's actually left. This is the single biggest lever on your bill — uploading both a transcript and a translation cuts the dubbing rate by up to 65%.

MP3MP4WAV M4AMOVCSVSRT
Stage 02
Transcribe

Audio becomes structured, editable text.

Expect 95%+ accuracy in primary languages — English, Spanish, French, German, Portuguese — with speaker identification built in. Other languages vary more, but still give you a working starting point rather than a blank page.

This is the stage that makes the rest of the pipeline hold together. The transcript becomes your master timeline, and every downstream step inherits its timecodes automatically. That's why dubs land in sync without anyone re-aligning them by hand. Before anything moves forward you can edit segments, split them, adjust timecodes, and assign or create speakers.

Speaker IDTimecode-lockedEditable
Stage 03
Translate

Your tone, across 80+ languages.

Pick a Translation Style and a Content Type — Cinema versus YouTube, Legal versus Dynamic — and the engine adapts voice, rhythm and terminology, not just words. A documentary and a gaming channel should not sound the same in Japanese, and they don't.

Context-aware adaptation reaches up to 90%, which mostly matters because it makes proofreading fast: you're correcting nuance, not rewriting meaning. Export as DOC, PDF, SRT or CSV to drop straight into any production workflow.

80+ languagesContext-awareDOC · PDF · SRT · CSV
Stage 04
Dub

AI dubbing, frame-perfect delivery.

Generate dubs in Standard or Deluxe — Standard for high-volume work where reliability matters most, Deluxe for audiences where premium voice quality isn't optional. Deliver final audio stems, or export video for QA review first.

Need the dub to sound like a specific person? Voice Cloning is available as an add-on at 8 credits per minute: upload an authorized voice sample once and every dub matches that tone across every language in the project.

Standard voicesDeluxe voices + Voice CloningVideo QA export

What each stage costs

Credits per minute of content. One balance covers every service — you're not buying four subscriptions. Uploading your own source files lowers the rate on the stages that come after them.

Service Full package With your files
Transcription 8
Translation Deluxe 18 14−22%
AI Dubbing · Standard 25 10−60%
AI Dubbing · Deluxe 37 20−46%
Voice Cloning add-on 8

Run the whole pipeline in one session.

Transcription, translation and dubbing on a single timeline, with one credit balance across all of them.

Get Started