AI Adversarial Specialist - Fully Remote | Upto $62/hr
About the job
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .
Position: AI Safety Experts โ English & Finnish
Type: Contract
Compensation: $48โ$62/hour
Location: Remote
Role Responsibilities
- Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structured approaches using taxonomies, benchmarks, and playbooks to ensure consistent testing.
- Document findings reproducibly to produce reports, datasets, and attack cases that customers can act on.
- Work independently and asynchronously to meet deadlines while improving AI model performance .
Qualifications
Must-Have
- Fluent in English and Finnish .
- Prior experience in red teaming, AI adversarial work , cybersecurity , or socio-technical probing.
- Ability to communicate risks clearly to both technical and non-technical stakeholders.
Preferred
- Experience with Adversarial ML , including jailbreak datasets, prompt injection, and model extraction.
- Background in Cybersecurity , such as penetration testing and exploit development.
- Expertise in socio-technical risk, including harassment/disinfo probing and abuse analysis.
Application Process (Takes 20โ30 mins to complete)
- Upload resume
- AI interview based on your resume
- Submit form
Resources & Support
- For details about the interview process and platform information, please check:
- For any help or support, reach out to:
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
Originally posted on Himalayas