AI Safety Expert - Red Teamer
RegionAustralia
Job description
About the job
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .
Position: AI Safety Experts — English & Vietnamese
Type: Contract
Compensation: $17–$25/hour
Location: Remote
Role Responsibilities
- Red team conversational AI models and agents by conducting jailbreaks, prompt injections, and bias exploitation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structure using taxonomies, benchmarks, and playbooks to ensure consistent testing.
- Document reproducibly by producing reports, datasets, and attack cases for customer action.
- Review AI outputs on sensitive topics like bias and misinformation, with optional participation in higher-sensitivity projects.
Qualifications
Must-Have
- Fluent in English and Vietnamese .
- Prior experience in red teaming, AI adversarial work , or cybersecurity .
- Ability to explain risks clearly to both technical and non-technical stakeholders.
Preferred
- Experience with adversarial ML , cybersecurity , and socio-technical risk.
- Skills in creative probing, psychology, acting, or writing for unconventional adversarial thinking.
Application Process (Takes 20–30 mins to complete)
- Upload resume
- AI interview based on your resume
- Submit form
Resources & Support
- For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
- For any help or support, reach out to: [email protected]
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
first seen 2026-08-12 14:12:01 · last verified 2026-08-22 20:30:01
pentestcareers.com // breach the job market