Live opening · Posted 1 day ago

AI Developer

AXS Solutions · Mumbai
Instahyre 3-6 yrs
You are 1 day behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 1 day ago
CompanyAXS Solutions
LocationMumbai
Experience3-6 yrs
SourceInstahyre
Listed1 day ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
18 min from Instahyre publishing this role to us finding it
14 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
16,981 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

We are seeking an experienced AI developer to lead the fine-tuning, deployment, and optimisation of the custom Proniti AI model based on the prevailing AI model architecture (26B-A4B MoE / 31B Dense). You will be responsible for transforming the base model into a highly secure, autonomous reasoning engine capable of executing complex standard operating procedure (SOP) gap analyses and regulatory reporting.
Responsibilities:
Model Fine-Tuning: Configure and execute supervised fine-tuning (SFT) pipelines using parameter-efficient fine-tuning (PEFT) methodologies. Utilise Quantised Low-Rank Adaptation (QLoRA) with frameworks like Hugging Face TRL and Unsloth (using bitsandbytes NF4 quantisation) to adapt the model without catastrophic forgetting.
Sovereign Infrastructure Deployment: Manage the deployment of the model on sovereign Indian cloud infrastructure. Work directly with dedicated infrastructure, NVIDIA H100 or L40S GPU clusters hosted in Mumbai-based Tier IV data centres to ensure data privacy and ultra-low latency.
Inference Optimisation: Deploy and configure the vLLM inference engine. You will optimise the server using flags like '--gpu-memory-utilisation' for long context management and enable Gemma 4's specific parsers (--reasoning-parser gemma4).
Agentic Tool Orchestration: Implement native tool-calling capabilities by mapping Proniti's backend APIs to the model's < |tool_call|> and < |tool_response|> control tokens, powering the autonomous reporting agent.
Constrained Decoding: Implement structured JSON output generation via vLLM's guided decoding engine to guarantee that the AI generates perfectly structured data payloads for the Proniti Compliance Dashboard.
Security Governance: Integrate the open-source Agent Governance Toolkit to provide deterministic, sub-millisecond policy enforcement, preventing risks like tool misuse or prompt injections.
Requirements:
3-5+ years of experience in deep learning, NLP, and AI systems engineering.
Strong proficiency in Python, PyTorch, and the Hugging Face ecosystem.
Proven hands-on experience with LLM/SLM fine-tuning techniques (LoRA, QLoRA) and quantisation.
Deep understanding of inference servers (specifically vLLM) and GPU memory optimisation (KV caching, PagedAttention).
Experience building autonomous AI agents and utilising JSON schemas for strict output decoding.
Qualifications: BE IT / BSc IT / MSc IT / MCA / ME IT / M. Tech IT or equivalent.

Experience
3-6 yrs

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App