Back to jobs

AI SAFETY RED TEAMER EXPERT

mercor
Part-timesenior€70-84/hour

Job description

About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Red Teamer Type: Contract Compensation: $70–$84/hour Location: Remote Role Responsibilities • Design adversarial prompts to stress-test frontier AI models . • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures. • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains. • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports. • Collaborate with AI researchers to improve model alignment, robustness, and safety. Qualifications Must-Have • Bachelor's degree or higher in Computer Science , Cybersecurity , Journalism , Communications , Psychology , Biology , Chemistry , Public Policy , or a related discipline. • 5+ years of professional experience in AI Safety , AI Red Teaming , Trust & Safety , cybersecurity, investigative journalism, life sciences, or a related field. • Strong analytical reasoning, prompt design, and written communication skills. • Experience designing adversarial prompts or evaluating frontier AI systems . Preferred • Experience with AI Red Teaming , RLHF , SFT , AI Alignment , or Trust & Safety . • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies. • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety. Application Process (Takes 20–30 mins to complete) • Upload resume • AI interview based on your resume • Submit form Resources & Support • For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome • For any help or support, reach out to: support@mercor.com PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

Skills

AI SafetyAI Red TeamingTrust & SafetyCybersecurityInvestigative JournalismLife SciencesPrompt EngineeringAdversarial EvaluationJailbreak TestingRLHFSFTAI Alignment