← All jobs

AI Safety Red Teamer (Train AI Models Part Time!)

up to $400k/year Remote Full-time
Intelligence AnalystResearch ScientistCompliance AnalystPrompt EngineerCyber Security ResearcherSecurity AnalystPenetration Tester

hackajob is partnering with Mercor to fill this position. Create a free profile and Archer will check you against this role and every other live role, showing you exactly where you match.

We are seeking experienced **AI Safety Red Teamers** to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area") topics. ## Responsibilities - Design adversarial prompts to stress-test frontier AI models. - Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures. - Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains. - Document vulnerabilities and contribute to safety benchmarking and red-teaming reports. - Collaborate with AI researchers to improve model alignment, robustness, and safety. ## Required Qualifications - Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline. - 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field. - Strong analytical reasoning, prompt design, and written communication skills. - Experience designing adversarial prompts or evaluating frontier AI systems. ## Preferred Qualifications - Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety. - Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies. - Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety. ## Why Join? - Help secure and strengthen the next generation of frontier AI models. - Work on cutting-edge adversarial testing alongside leading AI researchers and safety teams. - Influence how AI systems respond to complex, real-world safety challenges.

Apply knowing you're qualified

One free profile is all it takes. Archer checks you against this role and every other live role on hackajob, and shows you exactly which requirements you meet before you apply.

More roles like this

See all matching roles

Not quite the right role?

Archer scans thousands of live roles and surfaces the ones you genuinely match, each with a clear explanation of why. It keeps working after you apply, so you hear about roles you would never have found by searching.

Create your free profile