🎁 Before you apply, rehearse this interview. Create your free WorkMundi account and get an Interview Training on HelpsYouSpeak — no cost, no card. I want my training →
Requirements 710 years of experience in ML/Speech Engineering Strong expertise in ASR (Whisper, Deepgram, Azure Speech, Google STT) Experience with LLMs, RAG & Fine-tuning Knowledge of SIP, WebRTC, Twilio, Exotel or Ozonetel Experience with Voicebots in production Ability to optimize latency and scale Voice AI systems Tech Lead- JD.docx Key Responsibilities Design end-to-end Voice AI architectureLead technical decisions and mentor engineersBuild scalable, production-ready Voice AI solutionsOptimize performance, latency, and accuracyPay: 1,000,000.00 - 4,000,000.00 per year Benefits: Cell phone reimbursementFlexible schedulePaid sick timeApplication Question(s): Machine Learning (ML) and Speech EngineeringVoice AI Architecture Automatic Speech Recognition (ASR) Whisper Deepgram Google Speech-to-Text Azure Speech Natural Language Understanding (NLU) Large Language Models (LLMs) Dialogue Management Systems Text-to-Speech (TTS) ElevenLabs Coqui Retrieval-Augmented Generation (RAG) LLM Fine-tuning Telephony Integration SIP WebRTC Twilio Exotel Ozonetel Latency Optimization Cost Optimization Production Voice Pipeline Development Voicebot Development & Deployment Speech Recognition Accuracy Optimization (WER) Intent Classification Dialect & Code-Switching Handling (Hindi/Hinglish, Gulf Arabic) Generative AI Production-scale AI Systems Technical Architecture & System Design Work Location: In person .