Live opening · Posted 5 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Voice AI Engineer
Experience: 3–6 Years
Location: Remote
Employment Type: Full-time
CTC: Up to ₹30 LPA
About the Role
We are looking for a Voice AI Engineer with 3–6 years of hands-on experience building and deploying voice-based AI solutions. You will work on conversational voice systems, real-time speech processing, and AI-powered voice agents used in production environments.
The ideal candidate should have strong experience across Speech-to-Text (STT), Text-to-Speech (TTS), LLMs, conversational AI, and real-time voice pipelines.
Key Responsibilities
Design, build, and deploy production-grade AI voice agents and conversational systems.
Develop real-time voice pipelines integrating STT, LLMs, and TTS.
Work with platforms and APIs such as OpenAI, Deepgram, ElevenLabs, Google Speech, Azure Speech, or similar technologies.
Build and optimize voice interactions for latency, accuracy, reliability, and natural conversation.
Integrate LLMs with voice applications and external APIs/tools.
Develop conversational flows, prompts, context management, and function/tool calling.
Work on real-time audio streaming, WebSockets, SIP/telephony integrations, and voice infrastructure.
Implement techniques to improve speech recognition, response quality, interruption handling, and conversational experience.
Monitor and troubleshoot production voice AI systems.
Collaborate with product and engineering teams to build scalable AI-powered voice solutions.
Required Skills
3–6 years of professional experience in Voice AI / Conversational AI / Speech AI.
Strong hands-on experience with STT, TTS, LLMs, and voice-agent development.
Experience building real-time voice applications or AI voice agents.
Strong programming skills in Python.
Experience with APIs, SDKs, REST APIs, and WebSockets.
Understanding of LLM architectures, prompt engineering, function calling, and conversational workflows.
Experience with at least one speech/voice platform such as Deepgram, ElevenLabs, Azure Speech, Google Speech, OpenAI, AssemblyAI, or similar.
Understanding of audio formats, streaming, latency optimization, and real-time communication.
Experience integrating AI systems with third-party APIs and backend services.
Good understanding of software engineering practices, debugging, testing, and deployment.
Good to Have
Experience with SIP, Twilio, WebRTC, telephony APIs, or contact-center technologies.
Experience building voice agents for customer support, sales, healthcare, finance, or other enterprise use cases.
Experience with RAG, vector databases, and AI agent frameworks.
Experience deploying AI applications on AWS, GCP, or Azure.
Familiarity with Docker and cloud-based deployments.
Experience optimizing voice AI systems for low latency and high concurrency.
What We Offer
Remote-first work environment
Opportunity to work on cutting-edge Voice AI and LLM applications
High ownership and impact
Competitive compensation of up to ₹30 LPA
Opportunity to work with modern AI and voice technologies
Work arrangement
Yes
More openings worth a look
Recently tracked roles with full details and direct application links.