Live opening · Posted 18 days ago

Lead AI Engineer (.NET/C#)

FNZ · Gurugram, Haryana, India (On-site)
Linkedin No
You are 18 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 18 days ago
CompanyFNZ
LocationGurugram, Haryana, India (On-site)
Work modeNo
SourceLinkedin
Listed18 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
8 min from Linkedin publishing this role to us finding it
7 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
73,002 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

AI Evaluations Team Lead
Location: Gurugram, India
Seniority: Senior (6+ years engineering experience)
Purpose: Lead the team responsible for building FNZ's AI evaluations framework across both technical and process dimensions, and drive implementation of that framework across AI solutions to ensure rigorous safety, performance, and compliance standards before production deployment.
Key Responsibilities:
Lead the team that defines, builds, and evolves FNZ's AI evaluations framework across both technical components and operating processes, aligned to FNZ's six-pillar framework (Task Performance, Safety & Compliance, Efficiency, Groundedness & Reasoning, Robustness, Suitability)
Establish evaluation standards, methodologies, tooling, and governance processes, and lead implementation of the framework across AI solutions by embedding it into FNZ's SDLC as mandatory release gates
Represent evaluations function in AI Governance Committee, providing risk assessments and release recommendations
Build, mentor, and lead a team of evaluation specialists responsible for developing the framework and partnering with AI solution teams to implement it consistently across the estate
Design and execute complex evaluations for high-risk AI agents; lead red teaming exercises for critical deployments
Communicate evaluation findings to technical and non-technical stakeholders; influence product roadmaps
Skills and Experience
6+ years of engineering experience, with 2+ years in AI/ML or security testing and 2+ years in evaluation.
Deep understanding of LLM-based agents, RAG architectures, and agentic AI systems (not model training)
Strong programming background with hands-on experience building evaluation tooling, harnesses, or automated assessment workflows for AI agents and solutions
Proven ability to design evaluation methodologies and frameworks for probabilistic AI systems, covering both technical measures and operational processes
Experience embedding evaluation, assurance, or control frameworks into software delivery lifecycles and release governance, ideally within regulated environments
Leadership experience building or managing cross-functional teams and driving adoption of standards across multiple AI products or solution teams
Ability to translate safety, risk, and compliance requirements into practical evaluation criteria, controls, and release recommendations for production AI solutions

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App