Live opening · Posted 9 hours ago

Lead Research Engineer, Data Quality

Clera · San Francisco
Ashby No FullTime
You are 9 hours behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 9 hours ago
CompanyClera
LocationSan Francisco
Job typeFullTime
Work modeNo
SkillsPython, Docker
SourceAshby
Listed9 hours ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
16 min from Ashby publishing this role to us finding it
16 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
71,560 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

ABOUT THE ROLE
This is a senior technical leadership role on the data quality team at an early-stage AI infrastructure company focused on building and scaling reinforcement learning environments for frontier model training. You will own the strategy and systems that measure and improve training data quality, shaping research culture around what makes agent data genuinely useful rather than just superficially correct.
WHAT YOU'LL DO
- Lead the data quality team in building systems that evaluate thousands of tasks across RL environments, synthetic data, benchmarks, and domain-specific workflows.
- Define the data quality strategy by building QC systems, enforcing standards, and designing experiments to grade agent outputs.
- Develop new methods for validating synthetic data at scale, including failure-mode analysis, task mutation checks, and trajectory auditing.
- Partner with research engineers, domain experts, and data vendors to diagnose quality issues and improve data generation workflows.
- Turn qualitative research insights into production systems: internal tools, dashboards, validation pipelines, and feedback loops.
- Mentor research engineers to maintain a high bar for technical rigor, clarity, and execution speed.
WHAT WE'RE LOOKING FOR
- 5+ years of relevant engineering or research experience, specifically building systems for AI/ML data evaluation or data quality.
- Proven track record leading technical teams on ambiguous projects from problem definition through implementation and iteration.
- Advanced proficiency in Python, Docker, and Linux environments.
- Experience building QC systems, evals, benchmarks, synthetic data pipelines, validation workflows, or model evaluation infrastructure.
- Deep intuition for what makes training tasks realistic, learnable, diverse, reliable, and useful for AI agents.
- Research-oriented understanding of AI evals and post-training, beyond surface-level agent tooling projects.
- Comfort designing metrics, experiments, and QA/QC processes, not just executing them.
- Strong written communication skills, with the ability to explain methodology clearly to diverse audiences.
- Experience working with subject-matter experts to capture domain judgment and convert it into scalable review or generation systems.
- Early-stage startup experience and the ability to work independently in fast-paced environments.
COMPENSATION & BENEFITS
Salary range: $150,000 to $180,000 USD annually. Visa sponsorship is available.
LOCATION
On-site in San Francisco, CA, USA. The team also has a presence in Singapore.

Employment type
FullTime

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App