Live opening · Posted 1 day ago

AI Safety Specialist - Bilingual

Mercor · Mumbai Metropolitan Region (Remote)
Linkedin Yes
You are 1 day behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 1 day ago
CompanyMercor
LocationMumbai Metropolitan Region (Remote)
Work modeYes
SourceLinkedin
Listed1 day ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
7 min from Linkedin publishing this role to us finding it
1 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
72,824 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

About The Job
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.
Position: AI Safety Experts — English & Malayalam
Type: Contract
Compensation: $16–$22/hour
Location: Remote
Role Responsibilities
Red team conversational AI models and agents to identify jailbreaks, prompt injections, and misuse cases.
Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
Apply structure by following taxonomies, benchmarks, and playbooks to ensure consistent testing.
Document reproducibly by producing reports, datasets, and attack cases that customers can act on.
Work independently and asynchronously to meet deadlines while improving AI model performance.
Qualifications
Must-Have
Fluent Language Skills Required: English & Malayalam. Native fluency in English and Malayalam is required.
Strong judgment about language and content accuracy.
Rigorous attention to detail and ability to notice subtle errors.
Structured approach to work following guidelines and quality standards.
Clear communication skills for technical and non-technical audiences.
Adaptability across projects, task types, and customers.
Preferred
Experience in Adversarial ML: jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction.
Cybersecurity skills: penetration testing, exploit development, reverse engineering.
Understanding of socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing.
Creative probing skills: psychology, acting, writing for unconventional adversarial thinking.
Application Process (Takes 20–30 mins to complete)
Upload resume
AI interview based on your resume
Submit form
Resources & Support
For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
For any help or support, reach out to: support@mercor.com
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

Work arrangement
Yes

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App