mercor

AI Safety Specialist - Fully Remote | Upto $22/hr

Malaysia · REMOTE · FREELANCE
Publiée le 21 septembre 2026 · Candidature traitée sur le site de l’entreprise
AI-SafetyRed-TeamingAI-ResearchContent-ModerationAI-TestingAI-Safety-SpecialistAI-Safety-AnalystAI-Safety-PractitionerSafety-AIAI-Safety-Evaluator

About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Tamil Type: Contract Compensation: $16–$22/hour Location: Remote Role Responsibilities • Red team conversational AI models and agents through jailbreaks, prompt injections, and misuse cases. Identify bias exploitation and multi-turn manipulation. • Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks. • Apply structure by following taxonomies, benchmarks, and playbooks to ensure consistent testing. • Document reproducibly by producing reports, datasets, and attack cases that customers can act on. • Work independently and asynchronously to meet deadlines while improving AI model performance . Qualifications Must-Have • Fluent Language Skills Required: English & Tamil. Native fluency in English and Tamil is required. • Strong judgment about language and content accuracy, completeness, and appropriateness. • Rigorous attention to subtle errors, inconsistencies, and gaps. • Structured approach to work, adhering to guidelines and quality standards. • Clear communication with both technical and non-technical audiences. • Adaptability across projects, task types, and customers. Preferred • Experience in Adversarial ML : jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction. • Cybersecurity skills: penetration testing, exploit development, reverse engineering. • Knowledge of socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing. • Creative probing skills: psychology, acting, writing for unconventional adversarial thinking. Application Process (Takes 20–30 mins to complete) • Upload resume • AI interview based on your resume • Submit form Resources & Support • For details about the interview process and platform information, please check: • For any help or support, reach out to: PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity. Originally posted on Himalayas