AI Safety Experts, English & Norwegian
$48 to $62/hIndicative range provided by Mercor
Posted on July 30, 2026 · Applications until October 18, 2026
I am sending you to the official Mercor page. Applying is free and in English.
I earn a commission from the platform when a candidate I referred gets hired. It changes nothing for you.
The AI Safety Expert red teams conversational AI models by using adversarial inputs to uncover vulnerabilities and generate safety data. The role requires native fluency in both English and Norwegian, alongside prior experience in AI red teaming, cybersecurity, or socio-technical probing.
Description in English, as published by Mercor.
Location: Remote
Fluent Language Skills Required: English & Norwegian. Native fluency in English and Norwegian is required for this position.
Why This Role Exists
At Mercor, we believe the safest AI is the one that’s already been attacked, by us. We are assembling a red team for this project - human data experts who probe AI models with adversarial inputs, surface vulnerabilities, and generate the red team data that makes AI safer for our customers.
This project involves reviewing AI outputs that touch on sensitive topics such as bias, misinformation, or harmful behaviors. All work is text-based, and participation in higher-sensitivity projects is optional and supported by clear guidelines and wellness resources. Before being exposed to any content, the topics will be clearly communicated.
What You’ll Do
-
Red team conversational AI models and agents: jailbreaks, prompt injections, misuse cases, bias exploitation, multi-turn manipulation
-
Generate high-quality human data: annotate failures, classify vulnerabilities, and flag systemic risks
-
Apply structure: follow taxonomies, benchmarks, and playbooks to keep testing consistent
-
Document reproducibly: produce reports, datasets, and attack cases customers can act on
Who You Are
-
You bring prior red teaming experience (AI adversarial work, cybersecurity, socio-technical probing)
-
You’re curious and adversarial: you instinctively push systems to breaking points
-
You’re structured: you use frameworks or benchmarks, not just random hacks
-
You’re communicative: you explain risks clearly to technical and non-technical stakeholders
-
You’re adaptable: thrive on moving across projects and customers
Nice-to-Have Specialties
-
Adversarial ML: jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction
-
Cybersecurity: penetration testing, exploit development, reverse engineering
-
Socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing
-
Creative probing: psychology, acting, writing for unconventional adversarial thinking
What Success Looks Like
-
You uncover vulnerabilities automated tests miss
-
You deliver reproducible artifacts that strengthen customer AI systems
-
Evaluation coverage expands: more scenarios tested, fewer surprises in production
-
Mercor customers trust the safety of their AI because you’ve already probed it like an adversary
Why Join Mercor
-
Build experience in human data-driven AI red teaming at the frontier of safety
-
Play a direct role in making AI systems more robust, safe, and trustworthy
30 openings
About Mercor
Mercor is a US marketplace hiring remote experts for AI and consulting projects, from law to engineering. The original posting is on their site.
View this job on MercorMore jobs in Data analysis
Senior Backend Engineer (Python, SQL & AI Integration)
This role builds Python backend services, APIs, and LLM powered automation workflows while managing PostgreSQL databases. The ideal candidate brings strong proficiency in Python and hands on experience with database query optimization.
- Python
- SQL
- GCP AppEngine
Posted yesterday
Subject Matter Expert, Chart & Data Visualization Analysis
The contractor interprets and analyzes domain-specific charts and data visualizations to design quantitative reasoning tasks and step-by-step solutions for AI training. Requirements include a bachelor degree in a relevant field and at least two years of experience working with quantitative data or charts.
- Chart interpretation accuracy
- Multi-step quantitative reasoning
- Question design / unambiguous task construction
- +2
Posted yesterday$25 to $50/h
iPhone Users: Earn Money for Sharing Your Fitness Data (US Only)
The contributor populates an Apple HealthKit profile with comprehensive fitness activity and links clinical records from a qualifying healthcare institution. Applicants must provide verified health data and clinical notes from an approved provider list.
Posted 2 days ago
Biostatistician
This remote biostatistician contractor designs expert evaluation tasks, authentic datasets, and grading rubrics to train AI systems. Requirements include an MS or PhD in biostatistics, statistics, or epidemiology, and four plus years of relevant experience.
- Methodological rigor
- Clinical interpretation
- Regulatory sensitivity
- +1
Posted 2 days ago$60 to $100/h