mercor
AI Adversarial Specialist - Fully Remote | Upto $62/hr
About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Norwegian Type: Contract Compensation: $48–$62/hour Location: Remote Role Responsibilities • Red team conversational AI models and agents. Focus on jailbreaks, prompt injections, misuse cases, and bias exploitation. • Generate high-quality human data. Annotate failures, classify vulnerabilities, and flag systemic risks. • Apply structure using taxonomies, benchmarks, and playbooks to maintain consistent testing. • Document reproducibly. Produce reports, datasets, and attack cases that customers can act on. • Work independently and asynchronously to meet deadlines while improving AI model performance . Qualifications Must-Have • Fluent in English & Norwegian . • Prior red teaming experience in AI adversarial work , cybersecurity , or socio-technical probing . • Strong communication skills to explain risks to technical and non-technical stakeholders. • Ability to adapt and thrive across various projects and customers. Preferred • Experience in Adversarial ML , Cybersecurity , or Socio-technical risk . • Skills in jailbreak datasets , prompt injection , RLHF/DPO attacks , or model extraction . Application Process (Takes 20–30 mins to complete) • Upload resume • AI interview based on your resume • Submit form Resources & Support • For details about the interview process and platform information, please check: • For any help or support, reach out to: PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity. Originally posted on Himalayas
