Back to job search
Mercor logo
MercorVerified Job Source

AI Safety Specialist - Fully Remote | Upto $22/hr

  • Toronto, Ontario, Canada
  • Remote
  • Posted Sep 1, 2026
  • 1 position

US$20–US$22 / hour

Opens an external site

Sign in to save this job
Employment type
Part-time
Experience level
Mid-level · 2+ years
Apply by
Sep 25, 2026
Posting language
English
Working hours
40 hours per week
Location requirements
Country, Toronto, Ontario, Canada
Seniority
Not Applicable

Job summary

Red team conversational AI models to identify vulnerabilities like jailbreaks and prompt injections. Generate high-quality human data by annotating failures and documenting findings to improve model performance.

Job details

About The Job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey. Position: AI Safety Experts — English & Assamese Type: Contract Compensation: $20–$22/hour Location: Remote Role Responsibilities Red team conversational AI models and agents to identify vulnerabilities such as jailbreaks, prompt injections, and misuse cases. Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks. Apply structure using taxonomies, benchmarks, and playbooks to ensure consistent testing. Document findings reproducibly to create reports, datasets, and attack cases for customer action. Work independently and asynchronously to meet deadlines while improving AI model performance. Qualifications Must-Have Fluent in English and Assamese. Prior experience in red teaming, AI adversarial work, or cybersecurity. Strong communication skills to explain risks to both technical and non-technical stakeholders. Preferred Experience with adversarial ML, cybersecurity, and socio-technical risk analysis. Skills in creative probing such as psychology, acting, or unconventional adversarial thinking. Application Process (Takes 20–30 mins to complete) Upload resume AI interview based on your resume Submit form Resources & Support For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome For any help or support, reach out to: support@mercor.com PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

What you’ll do

Red team conversational AI models to identify vulnerabilities like jailbreaks and prompt injections. Generate high-quality human data by annotating failures and documenting findings to improve model performance.

Requirements

Must be fluent in English and Assamese with prior experience in red teaming, AI adversarial work, or cybersecurity. Strong communication skills and experience with adversarial ML or socio-technical risk analysis are preferred.

Listed skills

  • Technical · Preferred
  • analysis · Preferred
  • safety · Preferred
  • Process · Preferred
  • Communication · Preferred
  • meet deadlines · Preferred
  • Communication Skills · Preferred
  • Consistent · Preferred
  • Strong communication skills · Preferred
  • Strong Communication · Preferred
  • Customer · Preferred
  • English · Preferred

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • Red Teaming
  • AI Adversarial Work
  • Cybersecurity
  • English Fluency
  • Assamese Fluency
  • Adversarial ML
  • Socio-technical Risk Analysis
  • Prompt Injection Identification
  • Data Annotation
  • Vulnerability Classification

Job areas

  • Security & Safety
  • Software
  • Technology
  • Data & Analytics
  • Science & Research

More jobs you can apply to directly

Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.

Browse all Easy Apply jobs