Requirements
✔ 7–10 years of experience in ML/Speech Engineering
✔ Strong expertise in ASR (Whisper, Deepgram, Azure Speech, Google STT)
✔ Experience with LLMs, RAG & Fine-tuning
✔ Knowledge of SIP, WebRTC, Twilio, Exotel or Ozonetel
✔ Experience with Voicebots in production
✔ Ability to optimize latency and scale Voice AI systems �
Tech Lead- JD.docx
Key Responsibilities
- Design end-to-end Voice AI architecture
- Lead technical decisions and mentor engineers
- Build scalable, production-ready Voice AI solutions
- Optimize performance, latency, and accuracy
Pay: ₹1,000,000.00 - ₹4,000,000.00 per year
Benefits:
- Cell phone reimbursement
- Flexible schedule
- Paid sick time
Application Question(s):
- Machine Learning (ML) and Speech Engineering
Voice AI Architecture
Automatic Speech Recognition (ASR)
Whisper
Deepgram
Google Speech-to-Text
Azure Speech
Natural Language Understanding (NLU)
Large Language Models (LLMs)
Dialogue Management Systems
Text-to-Speech (TTS)
ElevenLabs
Coqui
Retrieval-Augmented Generation (RAG)
LLM Fine-tuning
Telephony Integration
SIP
WebRTC
Twilio
Exotel
Ozonetel
Latency Optimization
Cost Optimization
Production Voice Pipeline Development
Voicebot Development & Deployment
Speech Recognition Accuracy Optimization (WER)
Intent Classification
Dialect & Code-Switching Handling (Hindi/Hinglish, Gulf Arabic)
Generative AI
Production-scale AI Systems
Technical Architecture & System Design
Work Location: In person