Join Sarvam AI as an Audio-Video Editor to refine AI-generated dubbing and localization outputs for Indian languages.
Posted by employer 2 days ago
First seen on Joblaze 9 hours ago
Last verified on the company career page 9 hours ago
Skills & Technologies
What you'll build
Must have
Nice to have
AI in the day-to-day
You'll work inside our platform, running videos through the pipeline and fixing what the model gets wrong.
Not disclosed in this posting: compensation, years of experience, visa sponsorship.
Joblaze summary
In this role, the Audio-Video Editor at Sarvam AI focuses on refining AI-generated dubbing and localization outputs, ensuring that the final audio aligns perfectly with the original video. Key skills include sound engineering fundamentals, familiarity with digital audio workstations, and a methodical approach to correcting timing, pacing, and mix levels. This position is suitable for individuals with experience in post-production audio, particularly those who can work efficiently with high volumes of files. The team emphasizes collaboration with engineering to enhance the dubbing pipeline, making it a hands-on role in a tech-driven environment.
Joblaze insights
Quick facts
From the original posting
Sarvam AI · Bengaluru/Hybrid · Full-time
Sarvam Studio builds AI systems that dub and localise video across Indian languages. The AI does the heavy lifting - transcription, translation, voice generation, timing. A human still has to make the final output sound and feel right.
That's this role. You'll work inside our platform, running videos through the pipeline and then fixing what the model gets wrong: timing drift, awkward breaths and pauses, mix levels that don't sit, lip-sync slips, artefacts in the generated voice. You're the last set of ears before a file ships to a client.
This is not a creative-director job and you don't need a showreel of brand films. We're looking for good technical ears and a methodical hand.
Process source video through our AI dubbing and editing pipeline, and review every output end-to-end
Correct generated audio: timing and sync against the original, pacing of dialogue, breaths, silences, emphasis
Mix and balance dubbed dialogue against original music and effects beds; ride levels, clean up noise, match loudness to delivery spec
Fix the video side where needed — cuts, sync trims, re-timing, subtitle burn-in, format and export specs
Flag recurring failure patterns back to the engineering team so the pipeline improves (this is a real part of the job, not an afterthought)
Maintain consistent quality standards and turnaround across a high volume of files
Good sound engineering fundamentals — you understand loudness standards, EQ, compression, noise reduction, dialogue editing and how a mix falls apart
Comfort with sync-critical work — matching audio to picture frame-accurately
An AI-first mindset: you're interested in working with a model's output rather than redoing it by hand, and you can tell the difference between "the model made a mistake" and "the model made a choice I wouldn't have"
Patience for volume work and consistency across many files
Experience in post-production audio mixing, dubbing, podcast editing, or similar. You don't need to be a senior mixer.
Experience with dubbing, ADR or localisation workflows
Working proficiency in a DAW (FL Studio, Pro Tools, Reaper, Audition, Logic) and at least one NLE (Premiere Pro, DaVinci Resolve, Final Cut)
Fluency or working comfort in one or more Indian languages beyond English
Familiarity with subtitle formats (SRT, VTT) and delivery specs for streaming platforms