Simpplr
Simpplr1d ago
New

Lead Voice AI Engineer

IndiaIndia·BangaloreHybridlead
Machine Learning EngineerData
1 views0 saves0 applied

Quick Summary

Key Responsibilities

ASR → LLM → tools → TTS. Evaluate and integrate leading speech, telephony, and AI technologies. Define architecture, engineering standards, and production readiness for the Voice AI platform.

Technical Tools
Machine Learning EngineerData

Simpplr is the AI-powered intranet for unifying the digital workplace. It brings people, trusted knowledge, apps, and agents into a coherent digital experience. Powered by a proprietary EX Knowledge Graph, Simpplr synthesizes signals and context across connected systems to deliver personalized information and actions. The platform serves as a digital hub supporting communications, engagement, employee services, and work. With low-code extensibility and enterprise-grade security and governance, Simpplr enables confident operation at scale. More than 1,000 organizations — including AAA, the NHS, Penske, and Moderna — trust Simpplr to keep their workforce informed, aligned, and productive. Learn more at simpplr.com.

About the Role

~1 min read

We are looking for a Lead Voice AI Engineer to build production-grade Voice Agents for frontline heavy verticals like healthcare, manufacturing, warehousing, retail, hospitality focusing on employee support, procurement, collections, logistics, ordering etc.

You will lead the design of low-latency, real-time voice systems combining ASR, TTS, LLMs, conversational AI, enterprise workflows, knowledge retrieval, compliance, and human handoff.

This is a hands-on technical leadership role for someone who can take Voice AI from architecture to production.

Responsibilities

~2 min read
  • Design and build the real-time voice runtime for live conversations.
  • Build and optimize streaming ASR, TTS, VAD, endpointing, turn-taking, and barge-in.
  • Build adaptive voice pipelines for high-noise frontline environments (60-112 dB), hospitals, factory floors, warehouses, including server-side noise cancellation, echo suppression, and dynamic ASR/TTS optimization for PSTN and mobile phone audio quality.
  • Architect multi-provider speech routing across a broad multilingual matrix, including code-switching (e.g., Spanglish, Hinglish), where no single ASR or TTS provider covers all languages, and language detection, provider selection, and fallback chains must operate in real time mid-call.
  • Develop Voice Agents that support multi-turn, multi-intent conversations, context switching, clarification, and recovery.
  • Integrate Voice Agents with workflows, APIs, CRM, ITSM, knowledge bases, and enterprise systems.
  • Build secure identity verification, consent, privacy, audit, and compliance controls.
  • Implement warm transfer, callback, queue routing, and seamless human handoff with full conversation context.
  • Optimize multilingual voice quality across accents, noisy environments, latency, and naturalness.
  • Build evaluation frameworks for WER, intent accuracy, response latency, containment, resolution, escalation, and CSAT.
  • Establish production observability across the full call path: ASR → LLM → tools → TTS.
  • Evaluate and integrate leading speech, telephony, and AI technologies.
  • Define architecture, engineering standards, and production readiness for the Voice AI platform.
  • Mentor engineers and lead critical technical design reviews.

 

Requirements

~1 min read
  • 7+ years of software engineering experience.
  • Strong experience building production distributed or real-time systems.
  • Hands-on experience with Conversational AI, Voice AI, Speech AI, or LLM-based agents.
  • Strong programming skills in Python, Java, Go, or equivalent.
  • Experience with APIs, streaming systems, asynchronous architectures, and cloud-native platforms.
  • Strong understanding of system design, scalability, reliability, and observability.

 

  • Hands-on experience with Deepgram for real-time ASR and streaming speech recognition.
  • Hands-on experience with LiveKit for WebRTC, real-time audio, voice-agent runtime, and session orchestration.
  • Hands-on experience with ElevenLabs for low-latency, natural TTS and conversational voice experiences.
  • Experience with OpenAI, Azure Speech, Google Speech, or similar ASR/TTS technologies.
  • Experience with WebRTC, SIP, RTP, WebSockets, Twilio, or contact-center platforms.
  • Experience with LLM agents, tool calling, RAG, LangGraph, or similar orchestration frameworks.
  • Experience integrating enterprise systems such as Salesforce, ServiceNow, Jira, Zendesk, or Workday.
  • Experience with multilingual speech, accent handling, noisy environments, PII redaction, and call-recording controls.
  • Experience building high-scale, multi-tenant SaaS platforms.

 

  • Voice Agents feel natural and responsive in real-time conversations.
  • Users can interrupt naturally and change context without breaking the conversation.
  • The system works reliably across languages, accents, and noisy environments.
  • Voice Agents securely execute enterprise workflows and grounded knowledge retrieval.
  • Complex cases escalate to humans with full context.
  • The platform meets measurable targets for latency, accuracy, reliability, containment, resolution, and customer satisfaction.

 

Voice is not chat with audio. Voice is a real-time interaction model with different requirements for latency, interruption, identity, compliance, failure handling, and human handoff.

 

We value the real you. To ensure a fair and authentic experience for everyone, we ask that you do not use AI tools (such as real-time answer generators, transcription apps, or note-taking bots) during your interview

Our process is designed to hear your unique story, thought process, and lived experience in real-time. Use of unauthorized AI tools may result in disqualification, as we want to ensure every candidate is evaluated on their own individual merits. We’re excited to meet the person behind the resume!

If you need assistive technology or AI tools for accessibility (e.g., live captioning), please notify your recruiter in advance. We are committed to providing an inclusive interview experience.

  • Hub - 100% work from Simpplr office. Role requires Simpplifier to be in the office full-time.
  • Hybrid - Hybrid work from home and office. Role dictates the ability to work from home, plus benefit from in-person collaboration on a regular basis. 
  • Remote - 100% remote. Role can be done anywhere within your country of hire, as long as the requirements of the role are met. 

Location & Eligibility

Where is the job
Bangalore, India
Hybrid — some on-site time required
Who can apply
IN

Listing Details

Posted
September 3, 2026
First seen
September 3, 2026
Last seen
September 3, 2026

Posting Health

Days active
0
Repost count
0
Trust Level
62%
Scored at
September 3, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Simpplr
Simpplr
greenhouse
Employees
350
Founded
2014
View company profile
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

SimpplrLead Voice AI Engineer