Opens an external site
- Employment type
- Full-time
- Experience level
- Senior · 5+ years
- Posting language
- English
- Working hours
- 40 hours per week
- Seniority
- Associate
- Application method
- Direct apply is available
Job summary
The role involves designing and maintaining Python code to train, optimize, and benchmark large language models. Responsibilities include creating high-quality datasets for supervised fine-tuning and collaborating on RLHF to refine reward models.
Job details
Role: Python Engineer (Remote) Location: Remote (Work from Anywhere) Role Overview: We are hiring for one of our clients, seeking a Senior Python Developer to assist a foundational LLM company in enhancing their large language models. The goal is to provide high-quality proprietary data for fine-tuning and benchmarking model performance. Key Responsibilities: • Design, develop, and maintain efficient, high-quality Python code to train and optimize AI models. • Conduct evaluations to benchmark model performance and analyze results for continuous improvement. • Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria. • Lead efforts in supervised fine-tuning, including creating and maintaining high-quality, task-specific datasets. • Collaborate with researchers and annotators to execute reinforcement learning with human feedback and refine reward models. Required Skills & Qualifications: • Proficiency in Python and related frameworks/libraries for AI/ML tasks. • Experience with supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF). • Strong understanding of evaluation strategies and benchmarking processes for AI models. • Ability to design and implement Python code for data generation and model optimization. • Familiarity with AI model response evaluation and ranking methodologies. More About the Opportunity: This role offers a unique opportunity to work with a global leader in artificial intelligence, contributing to the advancement of large language models. Candidates will collaborate with top researchers and engineers in the field. Equal Opportunity Employer: We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications. Apply Now!
What you’ll do
The role involves designing and maintaining Python code to train, optimize, and benchmark large language models. Responsibilities include creating high-quality datasets for supervised fine-tuning and collaborating on RLHF to refine reward models.
Requirements
Candidates must be proficient in Python and AI/ML frameworks with specific experience in SFT and RLHF. A strong understanding of AI model evaluation strategies and ranking methodologies is required.
Listed skills
- Python · Preferred
Other relevant skills
Identified from the job description. Confirm important requirements above.
- Python
- Supervised Fine-Tuning
- Reinforcement Learning With Human Feedback
- AI Model Benchmarking
- Data Generation
- Model Optimization
- LLM Evaluation
- AI/ML Frameworks
Job areas
- Software
- Technology
- Engineering
- Science & Research
- Data & Analytics
More jobs you can apply to directly
Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.
BC Public Schools
Manager, Financial Planning and Analysis
SponsoredDirect employerEasy Apply- On-site
- Victoria, BC
- Posted Sep 18, 2026
Sustainable Projects Group
Sustainability consultant/ Team Lead, Energy Consulting NOC 41400
SponsoredDirect employerEasy Apply- Hybrid
- vancouver v5l 4s1, BC
- Posted Sep 17, 2026
Bédard Ressources Humaines
ITAD Services Representative #1265
SponsoredDirect employerEasy Apply- On-site
- Mississauga, ON
- Posted Sep 16, 2026
