About the role
Role: Software Engineer – AI Evaluation (Remote) Location: Remote (Work from Anywhere) Job Type: Part-Time Payout: Competitive, based on experience
Role Overview: We are hiring for one of our clients, seeking a Senior Software Engineer – LLM Evaluation (US/Canada/WEU based) to work on a part-time basis. You will create cutting-edge datasets for training, benchmarking, and advancing large language models while collaborating closely with researchers. This includes curating code examples, refining AI-generated code, and designing verification mechanisms for software engineering tasks.
Key Responsibilities: • Curate code examples, provide precise solutions, and make corrections in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go. • Evaluate and refine AI-generated code for efficiency, scalability, and reliability. • Collaborate with cross-functional teams to enhance enterprise-level AI-driven coding solutions against industry benchmarks. • Build verification agents to assess code quality and identify error patterns. • Hypothesize on software engineering lifecycle steps and evaluate model capabilities against them.
Required Skills & Qualifications: • Several years of software engineering experience (3 years or more). • Strong expertise in building full-stack applications and deploying scalable, production-grade software. • Deep understanding of software engineering principles including prototyping, architecture design, API design, production implementation, launch, experiments, monitoring, and operational maintenance. • Proficiency in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go. • Experience designing verification mechanisms for automated solution validation.
More About the Opportunity: This role offers a unique opportunity to work with a global leader in the Technology, Information and Internet | IT Services and IT Consulting | Data Infrastructure and Analytics | Research Services industry, contributing to the advancement of frontier AI systems. The position supports customers in accelerating research and transforming AI proofs of concept into proprietary intelligence.
Equal Opportunity Employer: We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications.
Apply Now!
About Hire Feed
HireFeed: A QuikHire product.
Earn in dollars. On your hours. From wherever.
HireFeed is the AI-curated feed of contract, gig, and AI training jobs that pay in USD, built for the people who actually want to work, not the ones writing job descriptions.
Why it exists: the highest-paying contract work in the world - RLHF, AI evaluation, domain-expert training, senior contract engineering, sits scattered across hundreds of ATS feeds and platform job boards. Most of it never makes it to Indeed or LinkedIn, and the roles that do are usually stale by the time you find them. The best opportunities are also the most ephemeral.
HireFeed pulls from 500+ verified sources, including Outlier, Mercor, Surge AI, Micro1, Turing, Toloka, Appen, and Remotasks, and refreshes every 60 seconds. The moment a role closes at the source, it falls off the feed.
What you'll find here:
- AI training, RLHF, and evaluation: $20–$80/hr
- Domain experts: medical, legal, finance, math (PhD)
- Senior contract engineering, paid weekly
- Multilingual annotation, creative writing training, AI research
What you won't find: pay-undisclosed listings, expired roles, "still accepting applications" lies, recruiter middlemen, or data resale. Pay is a hard requirement. Apply links 302-redirect to the source ATS. We never see your résumé.
Free for candidates, forever. Funded by the partner side.
→ hirefeed.co.in
Similar Jobs
About the role
Role: Software Engineer – AI Evaluation (Remote) Location: Remote (Work from Anywhere) Job Type: Part-Time Payout: Competitive, based on experience
Role Overview: We are hiring for one of our clients, seeking a Senior Software Engineer – LLM Evaluation (US/Canada/WEU based) to work on a part-time basis. You will create cutting-edge datasets for training, benchmarking, and advancing large language models while collaborating closely with researchers. This includes curating code examples, refining AI-generated code, and designing verification mechanisms for software engineering tasks.
Key Responsibilities: • Curate code examples, provide precise solutions, and make corrections in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go. • Evaluate and refine AI-generated code for efficiency, scalability, and reliability. • Collaborate with cross-functional teams to enhance enterprise-level AI-driven coding solutions against industry benchmarks. • Build verification agents to assess code quality and identify error patterns. • Hypothesize on software engineering lifecycle steps and evaluate model capabilities against them.
Required Skills & Qualifications: • Several years of software engineering experience (3 years or more). • Strong expertise in building full-stack applications and deploying scalable, production-grade software. • Deep understanding of software engineering principles including prototyping, architecture design, API design, production implementation, launch, experiments, monitoring, and operational maintenance. • Proficiency in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go. • Experience designing verification mechanisms for automated solution validation.
More About the Opportunity: This role offers a unique opportunity to work with a global leader in the Technology, Information and Internet | IT Services and IT Consulting | Data Infrastructure and Analytics | Research Services industry, contributing to the advancement of frontier AI systems. The position supports customers in accelerating research and transforming AI proofs of concept into proprietary intelligence.
Equal Opportunity Employer: We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications.
Apply Now!
About Hire Feed
HireFeed: A QuikHire product.
Earn in dollars. On your hours. From wherever.
HireFeed is the AI-curated feed of contract, gig, and AI training jobs that pay in USD, built for the people who actually want to work, not the ones writing job descriptions.
Why it exists: the highest-paying contract work in the world - RLHF, AI evaluation, domain-expert training, senior contract engineering, sits scattered across hundreds of ATS feeds and platform job boards. Most of it never makes it to Indeed or LinkedIn, and the roles that do are usually stale by the time you find them. The best opportunities are also the most ephemeral.
HireFeed pulls from 500+ verified sources, including Outlier, Mercor, Surge AI, Micro1, Turing, Toloka, Appen, and Remotasks, and refreshes every 60 seconds. The moment a role closes at the source, it falls off the feed.
What you'll find here:
- AI training, RLHF, and evaluation: $20–$80/hr
- Domain experts: medical, legal, finance, math (PhD)
- Senior contract engineering, paid weekly
- Multilingual annotation, creative writing training, AI research
What you won't find: pay-undisclosed listings, expired roles, "still accepting applications" lies, recruiter middlemen, or data resale. Pay is a hard requirement. Apply links 302-redirect to the source ATS. We never see your résumé.
Free for candidates, forever. Funded by the partner side.
→ hirefeed.co.in