AI Safety Red Teamer
We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing.
Hiring companyMercor
About the role
You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area") topics.
Why Join?
- Help secure and strengthen the next generation of frontier AI models.
- Work on cutting-edge adversarial testing alongside leading AI researchers and safety teams.
- Influence how AI systems respond to complex, real-world safety challenges.
Contract and Payment Terms
- You will be engaged as an independent contractor.
- This is a fully remote role that can be completed on your own schedule.
- Projects can be extended, shortened, or concluded early depending on needs and performance.
- Your work at Mercor will not involve access to confidential or proprietary information from any employer, client, or institution.
Responsibilities
- Design adversarial prompts to stress-test frontier AI models.
- Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
- Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
- Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
- Collaborate with AI researchers to improve model alignment, robustness, and safety.
Required skills
- Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
- 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
- Strong analytical reasoning, prompt design, and written communication skills.
- Experience designing adversarial prompts or evaluating frontier AI systems.
Preferred skills
- Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
- Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
- Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.
Eligibility
Work arrangement: Fully remote
Applicant eligibility: 40 eligible countries
View all 40 eligible countries
Eligible countries: Albania, Austria, Belgium, Bosnia & Herzegovina, Bulgaria, Croatia, Czechia, Denmark, Estonia, Finland, France, Germany, Greece, Hungary, Iceland, Ireland, Italy, Kosovo, Latvia, Liechtenstein, Lithuania, Luxembourg, Malta, Moldova, Monaco, Netherlands, North Macedonia, Norway, Poland, Portugal, Romania, San Marino, Serbia, Slovakia, Slovenia, Spain, Sweden, Switzerland, United Kingdom, United States
Please note: We are unable to support H1-B or STEM OPT candidates at this time.
Compensation details
$70–$84 per hour.
- Payments are weekly on Stripe or Wise based on services rendered.
More about the hiring companyMercor

