Job at a glance
Lead Voice AI Engineer This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Lead Voice AI Engineer based in India. This is a hands-on technical leadership role focused on building production-grade Voice AI systems for frontline industries such as healthcare, manufacturing, retail, hospitality, warehousing, and logistics. You will lead the architecture and development of low-latency, real-time voice experiences that combine speech recognition, text-to-speech, LLMs, conversational AI, knowledge retrieval, and enterprise workflows.
The role requires solving complex challenges around interruptions, turn-taking, multilingual conversations, noisy environments, identity, compliance, and human handoff. You will take voice systems from architecture through production while establishing engineering standards for scalability, reliability, and observability. You will also evaluate emerging speech and AI technologies and integrate them into a high-performance voice platform.
As a technical leader, you will mentor engineers and guide critical design decisions while remaining deeply involved in implementation. This is an opportunity to shape the next generation of enterprise Voice AI experiences where real-time performance and natural human interaction are essential. apply for this job Accountabilities: Design, architect, and build real-time voice runtimes capable of supporting live, natural conversations at production scale.
Develop and optimize streaming ASR, TTS, voice activity detection, endpointing, turn-taking, interruption handling, and barge-in capabilities. Build adaptive voice pipelines for high-noise environments such as hospitals, factories, warehouses, and frontline workplaces, including noise cancellation, echo suppression, and dynamic speech optimization. Architect multi-provider speech systems capable of supporting multilingual conversations, accents, and code-switching, including real-time language detection, provider selection, and fallback strategies.
Develop Voice AI agents capable of handling multi-turn and multi-intent conversations, context switching, clarification, recovery, and complex user interactions. Integrate voice agents with enterprise workflows, APIs, CRM platforms, ITSM systems, knowledge bases, and other business applications. Implement secure identity verification, consent management, privacy controls, audit capabilities, and compliance mechanisms for voice interactions.
Build reliable human handoff capabilities, including warm transfers, callbacks, queue routing, and transfer of complete conversation context. Optimize voice experiences for latency, speech quality, multilingual performance, accent recognition, naturalness, and reliability. Establish evaluation frameworks covering metrics such as word error rate, intent accuracy, response latency, containment, resolution, escalation, and customer satisfaction.
Implement end-to-end observability across the voice interaction lifecycle, from ASR and LLM processing through tools, workflows, and TTS. Evaluate and integrate leading speech, telephony, conversational AI, and voice technologies. Define architecture principles, engineering standards, reliability requirements, and production-readiness criteria for the Voice AI platform. Lead critical technical design reviews and mentor engineers working on voice and conversational AI systems.
Drive continuous improvements in system scalability, performance, reliability, security, and user experience. Requirements 7+ years of professional software engineering experience. Strong track record of building and operating production-grade distributed, real-time, or highly scalable systems. Hands-on experience developing Conversational AI, Voice AI, Speech AI, or LLM-based agent systems. Strong programming expertise in Python, Java, Go, or an equivalent programming language.
Experience with APIs, streaming architectures, asynchronous systems, and cl