Tech Lead / Voice AI Architect
Location: Hyderabad (primary hub — deepest voice-AI talent pool in India outside Bangalore)
Experience: 7–10 years in ML/speech engineering, with 2+ years architecting production voice pipelines
Responsibilities
- Own end-to-end architecture: ASR → NLU/LLM → dialogue manager → TTS → telephony
- Make build-vs-buy calls on ASR/TTS vendors (Deepgram, Azure Speech, ElevenLabs, or open-source Whisper/Coqui)
- Own latency budget across the pipeline (target sub-1.5s round-trip for live calls)
- Mentor Conversational AI and ML/Voice engineers; own technical hiring bar
Qualifications
- Must-Have Technical Qualifications
- Voice AI & Speech Technologies
- Production experience with at least one ASR engine (Whisper, Deepgram, Google STT, Azure Speech) at scale
- Comfortable working across Hindi/Hinglish and English accent handling for Indian/UAE markets
- Experience handling dialect variation and code-switching (e.g. Hinglish, Gulf Arabic variants) in production ASR/NLU
- LLM & Conversational AI
- Hands-on experience with LLM-based dialogue systems (function calling, RAG, or fine-tuning)
- Telephony & Infrastructure
- Strong grasp of telephony integration: SIP, WebRTC, Twilio/Exotel/Ozonetel or similar
- Performance & Optimization
- Experience optimizing for latency and cost simultaneously in a live voice pipeline
- Track record of hitting production accuracy benchmarks: 95%+ transcription (WER-based) accuracy and 90%+ intent classification accuracy at scale — not just in a demo
Required Skills
- Good-to-Have
- Prior experience at Amazon Alexa, Microsoft Cortana/Speech, Google Assistant, or a voice-AI startup
- Exposure to Arabic ASR/TTS for UAE market
- 2+ years hands-on with LLMs/Generative AI in a production (not POC) setting