Tech Lead — ASR / TTS / Speech LLM (IC + Mentor)
Team: Gen AI
Location: Bengaluru
Commitment: Full-time
Workplace Type: onsite
What You’ll Do
* Prepare and maintain synthetic and real training datasets. * STT/TTS/Speech LLM and LLM training: model selection → fine-tuning → evaluation → deployment. * Build evaluation for clinical applications (RPM, triage, inbound/outbound). * Build scripts for data selection, augmentation (noise, codec, jitter), and corpus curation. * Fine-tune speech models using CTC/RNN-T or adapter-based recipes on multi-GPU systems. * Fine-tune LLMs using PEFT (LoRA/QLoRA/adapters) and preference methods such as DPO/RLHF where needed. * Implement evaluation pipelines to measure WER/sWER, entity F1, safety/quality metrics, and latency; automate MLflow logging. * Experiment with bias-aware training and context list conditioning. * Collaborate with backend and DevOps teams to integrate trained models into inference stacks. * Support creation of context biasing APIs and LM rescoring paths. * Assist in maintaining benchmarks versus commercial baselines (Deepgram, Elevenlabs, Cartesia, Whisper, etc.). * Optimize inference latency/cost for speech and LLM serving (batching, kv-cache, quantization, caching, autoscaling).
Desired Skills
* Strong programming in Python (PyTorch, Hugging Face, NeMo, ESPnet). * Practical experience in audio data processing, augmentation, and ASR fine-tuning. * Training: SpecAugment, speed perturb, noise/RIRs, codec+PLC+jitter sims for PSTN/WebRTC. * Streaming ASR: Transducer/zipformer with chunked attention, frame-sync beam search, endpointing (VAD-EOU) tuning. * Context biasing: WFST boosts + neural re-scoring; patient/name dictionaries; session-aware bias refresh. * Familiarity with LoRA/QLoRA/adapters, distributed training, mixed precision. * Experience with LLM alignment and evaluation (SFT, DPO/RLHF, tool calling reliability, hallucination/safety checks). * Proficiency with evaluation frameworks: WER/sWER, Entity-F1, DER/JER, MOSNet/BVCC (TTS), PESQ/STOI (telephony), RTF/latency at P95/P99, and MLflow logging. * Inference/serving familiarity: vLLM/Triton, quantization, kv-cache, batching, and performance tuning. * Frameworks: ESPnet, SpeechBrain, NeMo, Kaldi/K2, Livekit, Pipecat, Dify. * Understanding of telephony speech characteristics, accents, and distortions. * Collaborative mindset for cross-functional work with ML-ops and QA.
Qualifications
- M.S. / Ph.D. in Computer Science, Speech Processing, or related field.
- 7–10 years of experience in applied ML, at least 3 in speech or multimodal AI.
- Track record of shipping production ASR/TTS models or inference systems at scale.
There are more than 50,000 engineering jobs:
Subscribe to membership and unlock all jobs
Engineering Jobs
60,000+ jobs from 4,500+ well-funded companies
Updated Daily
New jobs are added every day as companies post them
Refined Search
Use filters like skill, location, etc to narrow results
Become a member
🥳🥳🥳 452 happy customers and counting...
Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.
To try it out
For active job seekers
For those who are passive looking
Cancel anytime
Frequently Asked Questions
- We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
- We've got over 200,000 jobs from 15,000+ vetted companies. No fake or sleazy jobs here!
- We aggregate jobs from 15,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
- We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
- Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
- Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
- Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅
What Fellow Engineers Say
