Live opening · Posted 13 days ago

Research Engineer, Multi-Domain Alignment (SLM)

Invyte.ai · Gandhinagar, Gujarat, India (On-site)
Linkedin No
You are 13 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 13 days ago
CompanyInvyte.ai
LocationGandhinagar, Gujarat, India (On-site)
Work modeNo
SourceLinkedin
Listed13 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
5 min from Linkedin publishing this role to us finding it
10 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
72,375 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

About Invyte, Inc
An AI-native hiring platform advancing Sovereign AI through compact, "right-sized" models that are steerable, reliable, and deployable across high-stakes domains.
Job Description
Bridge the gap between base pretraining and real-world deployment by architecting behavioral logic and safety frameworks for a new class of multimodal SLMs. Focus on multi-domain alignment to ensure models can transition seamlessly between specialized fields—Legal, Healthcare, and Industrial Robotics—while maintaining rigorous adherence to human intent and cultural values.
Key Responsibilities
Design and implement scalable alignment pipelines (SFT, DPO, PPO) to optimize 1B–7B parameter models for high-stakes, domain-specific tasks
Architect reward models and preference datasets that capture nuanced domain expertise, moving beyond generic helpfulness to expert-level reasoning
Develop innovative techniques to mitigate alignment drift and catastrophic forgetting when models are specialized across disparate industries
Devise rigorous, automated benchmarking suites (LLM-as-a-judge) and adversarial testing frameworks to validate model robustness in out-of-distribution scenarios
Contribute to the broader AI community by open-sourcing high-quality code and producing reproducible research
Qualifications
Master's or PhD in Computer Science, ML, or equivalent practical experience in training large-scale models
Expertise in Python and PyTorch, specifically within the Hugging Face ecosystem (Transformers, TRL, PEFT, Accelerate)
Significant experience with RLHF, Direct Preference Optimization (DPO), and Constitutional AI
Deep understanding of Scaling Laws and the Alignment Tax—maximizing performance in compute-constrained environments
Experience aligning models that process text, visual, and sensor-based data
Research results published at leading venues such as NeurIPS, ICML, ICLR, or MLSys (bonus)
Experience building high-fidelity synthetic data pipelines to improve multi-step reasoning and logic (bonus)
Familiarity with optimizing inference engines (vLLM, TensorRT-LLM) or writing custom kernels (Triton/CUDA) for edge deployment (bonus)

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App