About the role
Role: Python Engineer (Remote)
Location: Remote (Work from Anywhere) Job Type: Full-Time Payout: Competitive, based on experience
Role Overview
We are hiring for one of our clients, seeking a Senior Python Developer to assist a foundational LLM company in enhancing their large language models. The goal is to provide high-quality proprietary data for fine-tuning and benchmarking model performance.
Key Responsibilities
- Design, develop, and maintain efficient, high-quality Python code to train and optimize AI models.
- Conduct evaluations to benchmark model performance and analyze results for continuous improvement.
- Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria.
- Lead efforts in supervised fine-tuning, including creating and maintaining high-quality, task-specific datasets.
- Collaborate with researchers and annotators to execute reinforcement learning with human feedback and refine reward models.
Required Skills & Qualifications
- Proficiency in Python and related frameworks/libraries for AI/ML tasks.
- Experience with supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF).
- Strong understanding of evaluation strategies and benchmarking processes for AI models.
- Ability to design and implement Python code for data generation and model optimization.
- Familiarity with AI model response evaluation and ranking methodologies.
More About the Opportunity
This role offers a unique opportunity to work with a global leader in artificial intelligence, contributing to the advancement of large language models. Candidates will collaborate with top researchers and engineers in the field.
Equal Opportunity Employer
We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications.
Apply Now!
Not the right fit? Search for Python Engineer jobs in Canada
About Hire Feed
HireFeed: A QuikHire product.
Earn in dollars. On your hours. From wherever.
HireFeed is the AI-curated feed of contract, gig, and AI training jobs that pay in USD, built for the people who actually want to work, not the ones writing job descriptions.
Why it exists: the highest-paying contract work in the world - RLHF, AI evaluation, domain-expert training, senior contract engineering, sits scattered across hundreds of ATS feeds and platform job boards. Most of it never makes it to Indeed or LinkedIn, and the roles that do are usually stale by the time you find them. The best opportunities are also the most ephemeral.
HireFeed pulls from 500+ verified sources, including Outlier, Mercor, Surge AI, Micro1, Turing, Toloka, Appen, and Remotasks, and refreshes every 60 seconds. The moment a role closes at the source, it falls off the feed.
What you'll find here:
- AI training, RLHF, and evaluation: $20–$80/hr
- Domain experts: medical, legal, finance, math (PhD)
- Senior contract engineering, paid weekly
- Multilingual annotation, creative writing training, AI research
What you won't find: pay-undisclosed listings, expired roles, "still accepting applications" lies, recruiter middlemen, or data resale. Pay is a hard requirement. Apply links 302-redirect to the source ATS. We never see your résumé.
Free for candidates, forever. Funded by the partner side.
→ hirefeed.co.in
Similar Jobs
About the role
Role: Python Engineer (Remote)
Location: Remote (Work from Anywhere) Job Type: Full-Time Payout: Competitive, based on experience
Role Overview
We are hiring for one of our clients, seeking a Senior Python Developer to assist a foundational LLM company in enhancing their large language models. The goal is to provide high-quality proprietary data for fine-tuning and benchmarking model performance.
Key Responsibilities
- Design, develop, and maintain efficient, high-quality Python code to train and optimize AI models.
- Conduct evaluations to benchmark model performance and analyze results for continuous improvement.
- Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria.
- Lead efforts in supervised fine-tuning, including creating and maintaining high-quality, task-specific datasets.
- Collaborate with researchers and annotators to execute reinforcement learning with human feedback and refine reward models.
Required Skills & Qualifications
- Proficiency in Python and related frameworks/libraries for AI/ML tasks.
- Experience with supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF).
- Strong understanding of evaluation strategies and benchmarking processes for AI models.
- Ability to design and implement Python code for data generation and model optimization.
- Familiarity with AI model response evaluation and ranking methodologies.
More About the Opportunity
This role offers a unique opportunity to work with a global leader in artificial intelligence, contributing to the advancement of large language models. Candidates will collaborate with top researchers and engineers in the field.
Equal Opportunity Employer
We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications.
Apply Now!
Not the right fit? Search for Python Engineer jobs in Canada
About Hire Feed
HireFeed: A QuikHire product.
Earn in dollars. On your hours. From wherever.
HireFeed is the AI-curated feed of contract, gig, and AI training jobs that pay in USD, built for the people who actually want to work, not the ones writing job descriptions.
Why it exists: the highest-paying contract work in the world - RLHF, AI evaluation, domain-expert training, senior contract engineering, sits scattered across hundreds of ATS feeds and platform job boards. Most of it never makes it to Indeed or LinkedIn, and the roles that do are usually stale by the time you find them. The best opportunities are also the most ephemeral.
HireFeed pulls from 500+ verified sources, including Outlier, Mercor, Surge AI, Micro1, Turing, Toloka, Appen, and Remotasks, and refreshes every 60 seconds. The moment a role closes at the source, it falls off the feed.
What you'll find here:
- AI training, RLHF, and evaluation: $20–$80/hr
- Domain experts: medical, legal, finance, math (PhD)
- Senior contract engineering, paid weekly
- Multilingual annotation, creative writing training, AI research
What you won't find: pay-undisclosed listings, expired roles, "still accepting applications" lies, recruiter middlemen, or data resale. Pay is a hard requirement. Apply links 302-redirect to the source ATS. We never see your résumé.
Free for candidates, forever. Funded by the partner side.
→ hirefeed.co.in