AI Safety Specialist - Fully Remote | Upto $22/hr
- Toronto, Ontario, Canada
- Remote
- Posted Sep 1, 2026
- 1 position
US$20–US$22 / hour
Opens an external site
- Employment type
- Part-time
- Experience level
- Mid-level · 2+ years
- Apply by
- Sep 25, 2026
- Posting language
- English
- Working hours
- 40 hours per week
- Location requirements
- Country, Toronto, Ontario, Canada
- Seniority
- Not Applicable
Job summary
Red team conversational AI models to identify vulnerabilities like jailbreaks and prompt injections. Generate high-quality human data by annotating failures and documenting findings to improve model performance.
Job details
About The Job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey. Position: AI Safety Experts — English & Assamese Type: Contract Compensation: $20–$22/hour Location: Remote Role Responsibilities Red team conversational AI models and agents to identify vulnerabilities such as jailbreaks, prompt injections, and misuse cases. Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks. Apply structure using taxonomies, benchmarks, and playbooks to ensure consistent testing. Document findings reproducibly to create reports, datasets, and attack cases for customer action. Work independently and asynchronously to meet deadlines while improving AI model performance. Qualifications Must-Have Fluent in English and Assamese. Prior experience in red teaming, AI adversarial work, or cybersecurity. Strong communication skills to explain risks to both technical and non-technical stakeholders. Preferred Experience with adversarial ML, cybersecurity, and socio-technical risk analysis. Skills in creative probing such as psychology, acting, or unconventional adversarial thinking. Application Process (Takes 20–30 mins to complete) Upload resume AI interview based on your resume Submit form Resources & Support For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome For any help or support, reach out to: support@mercor.com PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
What you’ll do
Red team conversational AI models to identify vulnerabilities like jailbreaks and prompt injections. Generate high-quality human data by annotating failures and documenting findings to improve model performance.
Requirements
Must be fluent in English and Assamese with prior experience in red teaming, AI adversarial work, or cybersecurity. Strong communication skills and experience with adversarial ML or socio-technical risk analysis are preferred.
Listed skills
- Technical · Preferred
- analysis · Preferred
- safety · Preferred
- Process · Preferred
- Communication · Preferred
- meet deadlines · Preferred
- Communication Skills · Preferred
- Consistent · Preferred
- Strong communication skills · Preferred
- Strong Communication · Preferred
- Customer · Preferred
- English · Preferred
Other relevant skills
Identified from the job description. Confirm important requirements above.
- Red Teaming
- AI Adversarial Work
- Cybersecurity
- English Fluency
- Assamese Fluency
- Adversarial ML
- Socio-technical Risk Analysis
- Prompt Injection Identification
- Data Annotation
- Vulnerability Classification
Job areas
- Security & Safety
- Software
- Technology
- Data & Analytics
- Science & Research
More jobs you can apply to directly
Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.
Bédard Ressources Humaines
Table Games Trainer #1414
SponsoredDirect employerEasy Apply- On-site
- Posted Sep 8, 2026
Trista
Residential Home Care Services Manager (NOC 60040)
SponsoredDirect employerEasy Apply- On-site
- Posted Sep 11, 2026
Bédard Ressources Humaines
Plant Manager - Food Industry #1812
SponsoredDirect employerEasy Apply- On-site
- Posted Sep 9, 2026
