

Mercor
Mercor
QA/Test Engineer
QA/Test Engineer

US

Contract-based

Date Posted

Offered salary
$60 - $90 per hour
$60 - $90 per hour

Closing date
Closing soon
Closing soon


Qualification
MSc / PhD in stem field
MSc / PhD in stem field


Hiring location
US
US


Experience
1+ Years
1+ Years
Responsibilities
• Design test cases that confirm each task works as intended, including tricky edge cases
• Review tasks and reference solutions before finalization, catching ambiguity and gaps early
• Debug tasks or checks in Python when they don’t behave as expected
• Help build simple, repeatable quality checklists and provide actionable feedback to authors
• Watch for shortcuts and grading gaps in AI agent runs to keep benchmark scores trustworthy
Requirements
• MSc or PhD in a STEM field, or equivalent practical experience in a research-heavy or engineering-heavy domain
• 1+ years of experience in test engineering, quality assurance, or a research/software engineering role with strong quality ownership
• Demonstrated skill designing test cases and quality-review processes, and debugging complex systems end-to-end
• Working proficiency in Python and Git, with comfort navigating unfamiliar codebases and environments
• Exceptional attention to detail and clear written documentation habits
• Perfectionist mindset with creativity in finding what others missed and ability to work independently on ambiguous problems
• Ability to engage reliably for ~35 hours per week
Preferred
• Past experience in AI training, model evaluation, or quality review of AI-generated work
How to Apply
Click "Apply" to be taken to the Mercor website. Complete your profile by uploading your resume and confirming your work location. Once verified, you will be matched to opportunities as they arise. Please note that this role cannot support H1B or STEM OPT candidates. Applying through our link supports WFH Bulletin as a referral partner, but you are welcome to apply directly if you prefer.
Responsibilities
• Design test cases that confirm each task works as intended, including tricky edge cases
• Review tasks and reference solutions before finalization, catching ambiguity and gaps early
• Debug tasks or checks in Python when they don’t behave as expected
• Help build simple, repeatable quality checklists and provide actionable feedback to authors
• Watch for shortcuts and grading gaps in AI agent runs to keep benchmark scores trustworthy
Requirements
• MSc or PhD in a STEM field, or equivalent practical experience in a research-heavy or engineering-heavy domain
• 1+ years of experience in test engineering, quality assurance, or a research/software engineering role with strong quality ownership
• Demonstrated skill designing test cases and quality-review processes, and debugging complex systems end-to-end
• Working proficiency in Python and Git, with comfort navigating unfamiliar codebases and environments
• Exceptional attention to detail and clear written documentation habits
• Perfectionist mindset with creativity in finding what others missed and ability to work independently on ambiguous problems
• Ability to engage reliably for ~35 hours per week
Preferred
• Past experience in AI training, model evaluation, or quality review of AI-generated work
How to Apply
Click "Apply" to be taken to the Mercor website. Complete your profile by uploading your resume and confirming your work location. Once verified, you will be matched to opportunities as they arise. Please note that this role cannot support H1B or STEM OPT candidates. Applying through our link supports WFH Bulletin as a referral partner, but you are welcome to apply directly if you prefer.
Responsibilities
• Design test cases that confirm each task works as intended, including tricky edge cases
• Review tasks and reference solutions before finalization, catching ambiguity and gaps early
• Debug tasks or checks in Python when they don’t behave as expected
• Help build simple, repeatable quality checklists and provide actionable feedback to authors
• Watch for shortcuts and grading gaps in AI agent runs to keep benchmark scores trustworthy
Requirements
• MSc or PhD in a STEM field, or equivalent practical experience in a research-heavy or engineering-heavy domain
• 1+ years of experience in test engineering, quality assurance, or a research/software engineering role with strong quality ownership
• Demonstrated skill designing test cases and quality-review processes, and debugging complex systems end-to-end
• Working proficiency in Python and Git, with comfort navigating unfamiliar codebases and environments
• Exceptional attention to detail and clear written documentation habits
• Perfectionist mindset with creativity in finding what others missed and ability to work independently on ambiguous problems
• Ability to engage reliably for ~35 hours per week
Preferred
• Past experience in AI training, model evaluation, or quality review of AI-generated work
How to Apply
Click "Apply" to be taken to the Mercor website. Complete your profile by uploading your resume and confirming your work location. Once verified, you will be matched to opportunities as they arise. Please note that this role cannot support H1B or STEM OPT candidates. Applying through our link supports WFH Bulletin as a referral partner, but you are welcome to apply directly if you prefer.


Mercor
QA/Test Engineer
QA/Test Engineer
Overview
Overview
Cincinnatus LLC is recruiting QA/Test Engineers for a leading AI lab to act as the quality backbone for complex multi-step benchmark tasks. You will design test cases, review tasks and reference solutions, debug issues, and help build repeatable quality processes. This is a full-time W-2 contingent role (~35 hours/week) with the opportunity to be placed at a leading AI lab.
Cincinnatus LLC is recruiting QA/Test Engineers for a leading AI lab to act as the quality backbone for complex multi-step benchmark tasks. You will design test cases, review tasks and reference solutions, debug issues, and help build repeatable quality processes. This is a full-time W-2 contingent role (~35 hours/week) with the opportunity to be placed at a leading AI lab.
Cincinnatus LLC is recruiting QA/Test Engineers for a leading AI lab to act as the quality backbone for complex multi-step benchmark tasks. You will design test cases, review tasks and reference solutions, debug issues, and help build repeatable quality processes. This is a full-time W-2 contingent role (~35 hours/week) with the opportunity to be placed at a leading AI lab.
Get Started
Find Verified Remote Jobs That Fit Your Career Goals
Explore carefully reviewed remote job opportunities from trusted companies worldwide. Discover roles that match your skills, experience and work preferences all in one place.
Newsletter
Get Started
Find Verified Remote Jobs That Fit Your Career Goals
Explore carefully reviewed remote job opportunities from trusted companies worldwide. Discover roles that match your skills, experience and work preferences all in one place.
Newsletter
Get Started
Find Verified Remote Jobs That Fit Your Career Goals
Explore carefully reviewed remote job opportunities from trusted companies worldwide. Discover roles that match your skills, experience and work preferences all in one place.

