AI Safety Expert - Red Teaming

Mercor·Toronto, Ontario·Posted 13d ago·via Talent.com

RegionCanada
SalaryCAD 48 - 62
Apply Now

Job description

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .

Position: AI Safety Experts — English & Swedish

Type: Contract

Compensation: $48–$62/hour

Location: Remote

Role Responsibilities

  • Red team conversational AI models and agents. Conduct jailbreaks, prompt injections, misuse cases, and bias exploitation.
  • Generate high-quality human data. Annotate failures, classify vulnerabilities, and flag systemic risks.
  • Apply structure using taxonomies, benchmarks, and playbooks to maintain consistent testing.
  • Document reproducibly by producing reports, datasets, and attack cases for customer action.
  • Work independently and asynchronously to meet deadlines while improving AI model performance .

Qualifications

Must-Have

  • Fluent in English and Swedish .
  • Prior red teaming experience in AI adversarial work , cybersecurity , or socio-technical probing.
  • Strong communication skills to explain risks to technical and non-technical stakeholders.

Preferred

  • Experience in Adversarial ML , Cybersecurity , or socio-technical risk analysis.
  • Skills in creative probing such as psychology, acting, or writing for unconventional adversarial thinking.

Application Process (Takes 20–30 mins to complete)

  • Upload resume
  • AI interview based on your resume
  • Submit form

Resources & Support

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

first seen 2026-08-12 08:24:01 · last verified 2026-08-23 17:00:01


pentestcareers.com // breach the job market

AI Safety Expert - Red Teaming at Mercor | Pentest Careers