AI video dubbing platform
Takes a video and dubs it into another language in a cloned voice, with optional lip sync and burned-in subtitles. Web, mobile and API.
How I built it
- Temporal workflows run each stage: transcription, translation, voice, stem separation, timeline placement, lip sync. A routing step sends single- and multi-speaker videos down different pipelines.
- Pipelines are locked and every output is stamped with its version. A job never switches provider halfway, so results stay consistent.
- Lip sync is refused on multi-speaker or uncertain footage instead of shipping a bad result.
- Top-ups are retry-safe and a worker recovers payments whose webhooks were missed. A per-minute cost model sits behind the pricing.
Taken from self-hosted GPU models to locked production pipelines, with about 50 test modules and a provider simulator for load tests that costs nothing to run.
- FastAPI
- Temporal
- Postgres
- Deepgram
- ElevenLabs
- Terraform
- Expo
