AI Safety Expert - Red Teaming

Mercor·San Francisco, California·Posted 12d ago·via Talent.com

RegionUSA
SalaryUSD 48 - 62
Apply Now

Job description

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .

Position: AI Safety Experts — English & Danish

Type: Contract

Compensation: $48–$62/hour

Location: Remote

Role Responsibilities

  • Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation.
  • Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
  • Apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing.
  • Document reproducibly by producing reports, datasets, and attack cases that customers can act on.
  • Work independently and asynchronously to meet deadlines while improving AI model performance .

Qualifications

Must-Have

  • Fluent in English and Danish .
  • Prior red teaming experience in AI adversarial work , cybersecurity , or socio-technical probing.
  • Strong communication skills to explain risks clearly to both technical and non-technical stakeholders.

Preferred

  • Experience in Adversarial ML , Cybersecurity , or socio-technical risk analysis.
  • Skills in creative probing such as psychology, acting, or writing for unconventional adversarial thinking.

Application Process (Takes 20–30 mins to complete)

  • Upload resume
  • AI interview based on your resume
  • Submit form

Resources & Support

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

first seen 2026-08-12 20:48:01 · last verified 2026-08-23 13:30:02


pentestcareers.com // breach the job market

AI Safety Expert - Red Teaming at Mercor | Pentest Careers