Live opening · Posted 7 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Preference modeling turns vague taste into something a system can optimize.
Through Nebulai's Humans & AI Agents Marketplace, you join teams that build preference models for enterprise AI: capture human feedback, learn ranking and reward signals, and keep models aligned under real product constraints.
You'll thrive here if you:
- Have built preference, reward, or ranking models for LLM or search products
- Care about feedback quality, bias, and stable optimization targets
- Partner with evaluation, safety, and product partners in enterprise settings
- Prefer grounded preference data over gut-feel tuning
Bonus if you've:
- Shipped RLHF, DPO, or pairwise preference pipelines at scale
- Supported a CoE standardizing preference bars across AI programs
- Tied weak preference signals to poor UX, churn, or trust scores
Contract work via Nebulai's marketplace. Apply at https://nebulai.app
Work arrangement
Yes
More openings worth a look
Recently tracked roles with full details and direct application links.