AI Safety Red Teamer stress-testing Mercor’s frontier AI models. Identifying vulnerabilities, jailbreaks, hallucinations, and policy failures through adversarial evaluation.
Responsibilities
Design adversarial prompts to stress-test frontier AI models
Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures
Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains
Document vulnerabilities and contribute to safety benchmarking and red-teaming reports
Collaborate with AI researchers to improve model alignment, robustness, and safety
Work on projects focused on training and enhancing AI systems
Requirements
Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline
5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field
Strong analytical reasoning, prompt design, and written communication skills
Experience designing adversarial prompts or evaluating frontier AI systems
Preferred: experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety
Preferred: familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies
Preferred: expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety
Must not require H1-B or STEM OPT sponsorship
Benefits
Fully remote role
Flexible schedule / work can be completed on your own schedule
Weekly payments via Stripe or Wise based on services rendered
Competitive pay
Opportunity to collaborate with leading AI researchers and safety teams
Opportunity to influence next-generation AI systems
Referral payments of up to $340 per successful referral
Senior AI Solutions Analyst guiding Desjardins Group’s financial services sectors in responsible productivity AI adoption. Assessing use cases, value, risks and Microsoft AI solutions.
AVP leading AI transformation, including fraud detection, for Manulife, a global financial services provider. Driving secure, governed Classical, Generative, and Agentic AI solutions across Corporate Functions.
Generative AI Analyst validating Canadian French translations for Welo Data, a global AI data company. Reviewing and annotating multilingual AI - generated content for training and quality improvement.
Senior engineers evaluating AI coding - agent interactions for G2i. Assessing reasoning, explanations, outputs, and engineering judgment across Codex, Claude Code, and Cursor.
Manager integrating workforce management with AI transformation at TD, a major North American bank. Connecting financial, operational, and workforce planning across customer operations.
Search +AI Discovery Manager leading enterprise SEO, AI - search strategy, and client advisory for dentsu, a global integrated marketing agency. Mentoring North American SEO teams and shaping innovative search methodologies.
Director leading Homebase’s data engineering, data science, and applied AI teams. Building trusted data infrastructure and AI - powered tools for 150,000+ small businesses.
AI Solution Specialist helping greenhouse growers adopt Source.ag’s AI - powered crop technology. Building dashboards, automations, and custom tools while translating greenhouse insights into product improvements.
Senior Advisor designing and integrating generative AI solutions for Desjardins, North America’s largest cooperative financial group. Driving adoption through training, change management and scalable AI practices.
Strategy and management consultant creating complex business scenarios and evaluating AI outputs. Supporting Gramian Consultancy’s IT talent solutions through rigorous consulting analysis and executive deliverables.