Live opening · Posted 1 day ago

Senior Generative AI Engineer

ConveGenius · Noida
Instahyre 5-9 yrs
You are 1 day behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 1 day ago
CompanyConveGenius
LocationNoida
Experience5-9 yrs
SourceInstahyre
Listed1 day ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
10 min from Instahyre publishing this role to us finding it
195 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
16,411 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Responsibilities:
Fine-tune foundation models for education-specific tasks: question answering, content generation, adaptive feedback, and curriculum alignment.
Own end-to-end fine-tuning workflows: dataset curation, training runs, hyperparameter tuning, evaluation, and model versioning.
Implement efficient fine-tuning methods (LoRA, QLoRA, DoRA, adapters) appropriate to available compute budgets.
Build RLHF and preference optimisation pipelines: DPO, PPO, reward modelling for aligning models to learning outcomes.
Optimise GPU training efficiency: DeepSpeed, FSDP, gradient checkpointing, mixed precision, multi-GPU setups.
Evaluate fine-tuned models rigorously: perplexity, task-specific benchmarks, human eval, and regression testing.
Build data pipelines for instruction tuning datasets, including multilingual and Indic language data.
Collaborate with AI platform teams on inference optimisation: quantisation, GGUF, ONNX export, vLLM serving.
Assess new open-source model releases for domain applicability and adoption readiness.
Define and track model performance metrics, evaluation benchmarks, and optimisation targets across all fine-tuned model versions.
Work with open-source and sovereign LLMs, owning the full model adaptation lifecycle using proven, industry-standard frameworks.
Requirements:
Strong experience in fine-tuning and optimising Large.
Hands-on experience with LoRA, QLoRA, SFT, DPO, Language Models (LLMs), RLHF, or similar fine-tuning techniques.
Proficiency in PyTorch and Hugging Face ecosystem (Transformers, PEFT, TRL).
Experience with distributed and multi-GPU model training.
Strong understanding of model performance, evaluation, and optimisation.
Knowledge of DeepSpeed, Megatron-LM, and large-scale training frameworks.
Understanding of AI/ML pipelines, data preparation, and model deployment.
Experience working with open-source and sovereign LLMs; focus on standard, production-proven frameworks.
Preferred Skills RAG integration: how fine-tuned models interact with retrieval systems and when to fine-tune vs retrieve.
Inference optimisation: GGUF quantisation, ONNX export, TensorRT, vLLM serving.
Vector databases: embedding generation, indexing, and hybrid search for downstream model use.
Multilingual fine-tuning: Indic languages, code-switching, transliteration-aware tokenisation.
Agentic AI: tool use, function calling, and fine-tuned model behaviour in agentic contexts.

Experience
5-9 yrs

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App