Live opening · Posted 4 days ago

Senior Voice AI / Speech ML Engineer

TalixoHR · Greater Bengaluru Area (On-site)
Linkedin No
You are 4 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 4 days ago
CompanyTalixoHR
LocationGreater Bengaluru Area (On-site)
Work modeNo
SourceLinkedin
Listed4 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
12 min from Linkedin publishing this role to us finding it
5 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
73,790 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Location: HSR Layout, Bengaluru | Work Mode: On-site | Experience: 3–4 Years | Employment: Full-time
About the Role
We are looking for a Senior Voice AI / Speech ML Engineer to build and train next-generation speech intelligence systems. This role is for an engineer who has hands-on experience training speech models and foundation models, not just integrating existing APIs or pretrained models.
What You’ll Work On
Develop, train, fine-tune and evaluate speech/foundation models for Voice AI.
Build and improve ASR, TTS, Speaker ID, VAD and conversational AI systems.
Work with large-scale speech, audio and text datasets.
Design training pipelines, experiments and evaluation frameworks.
Optimize models for accuracy, latency, robustness and inference efficiency.
Work with PyTorch/TensorFlow, Transformers and GPU-based training.
Deploy and optimize models for production environments.
Work on multilingual and Indian-language speech AI use cases.
Must-Have
3–4 years of relevant post-graduation/full-time experience in ML, Speech AI, NLP or related fields.
Personally trained/pre-trained foundation models — not only fine-tuned or consumed existing models.
Hands-on experience developing/training ASR or TTS models.
Strong Python + PyTorch/TensorFlow.
Experience with Transformers, GPUs and large-scale model training.
Strong understanding of model evaluation, optimization and debugging.
Good to Have
Indian/multilingual speech experience.
Whisper, wav2vec 2.0, HuBERT, NeMo, SpeechBrain or Hugging Face.
Distributed training, quantization and production deployment.
Audio preprocessing, augmentation, annotation and dataset quality.
Not a Fit If
Only GenAI/RAG/LLM application development.
Only prompt engineering or API integration.
Only using pretrained Whisper/TTS models.
Only fine-tuning models without foundation-model training experience.
Less than 3 years relevant post-graduation/full-time experience.

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App