Back to job search
Braintrust logo
BraintrustVerified Job Source

STEM PhD Expert for AI Reasoning & Evaluation

  • Argentina, Australia, Canada, Mexico, New Zealand, Puerto Rico, United States, United Kingdom
  • Remote
  • Posted Aug 28, 2026
  • 1 position

US$150 / hour

Opens an external site

Sign in to save this job
Employment type
Full-time
Experience level
Mid-level · 2+ years
Minimum education
Master’s degree
Posting language
English
Working hours
40 hours per week
Seniority
Mid-Senior level

Job summary

Review and evaluate AI-generated content to ensure factuality and relevance in STEM domains. Help improve the quality of technical reasoning by crafting questions and ranking model responses.

Job details

We're hiring PhD-level experts to help train and evaluate advanced AI models. This is remote, flexible contract work where your academic expertise directly shapes how cutting-edge models reason, answer, and improve. You'll review AI-generated content in your field, identify where the model gets things right or wrong, and help raise the quality bar on technical reasoning. This is a short-term engagement running through the end of June, with potential to extend depending on project needs. Key Responsibilities You May Contribute Your Expertise By Assessing the factuality and relevance of domain-specific text produced by AI models Crafting and answering questions related to Machine Learning and AI Evaluating and ranking domain-specific responses generated by AI models What We're Looking For A PhD (completed or in final stages) in one of the following fields: Machine Learning / AI, Computer Science, Engineering, Statistics, or a closely related quantitative subdomain such as Mathematics or Physics Strong analytical and critical-thinking skills, with the ability to spot subtle errors in technical reasoning Fluent written English and the ability to communicate complex ideas clearly Nice to Have Research experience (academic or industry) Prior experience with data annotation or AI model evaluation Experience reviewing or publishing research papers Compensation Up to $150/hr, depending on your area of expertise, depth of experience, and assessment performance. Location This role is fully remote. We are currently accepting applicants based in: United States, Canada, Puerto Rico, Mexico, United Kingdom, Australia, New Zealand, and Argentina.

What you’ll do

Review and evaluate AI-generated content to ensure factuality and relevance in STEM domains. Help improve the quality of technical reasoning by crafting questions and ranking model responses.

Requirements

Requires a PhD in Machine Learning, AI, Computer Science, Engineering, Statistics, Mathematics, or Physics. Must possess strong analytical skills and fluency in written English.

Listed skills

  • Machine learning · Preferred
  • Critical Thinking · Preferred

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • AI Model Evaluation
  • Technical Reasoning
  • Machine Learning
  • Data Annotation
  • Analytical Thinking
  • Critical Thinking
  • Technical Writing
  • Research

Job areas

  • Science & Research
  • Technology
  • Software
  • Data & Analytics
  • Engineering

More jobs you can apply to directly

Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.

Browse all Easy Apply jobs