Mercor

Mercor

QA/Test Engineer

QA/Test Engineer

US

Contract-based

Date Posted

Offered salary

$60 - $90 per hour

$60 - $90 per hour

Closing date

Closing soon

Closing soon

Qualification

MSc / PhD in stem field

MSc / PhD in stem field

Hiring location

US

US

Experience

1+ Years

1+ Years

Responsibilities
• Design test cases that confirm each task works as intended, including tricky edge cases
• Review tasks and reference solutions before finalization, catching ambiguity and gaps early
• Debug tasks or checks in Python when they don’t behave as expected
• Help build simple, repeatable quality checklists and provide actionable feedback to authors
• Watch for shortcuts and grading gaps in AI agent runs to keep benchmark scores trustworthy

Requirements
• MSc or PhD in a STEM field, or equivalent practical experience in a research-heavy or engineering-heavy domain
• 1+ years of experience in test engineering, quality assurance, or a research/software engineering role with strong quality ownership
• Demonstrated skill designing test cases and quality-review processes, and debugging complex systems end-to-end
• Working proficiency in Python and Git, with comfort navigating unfamiliar codebases and environments
• Exceptional attention to detail and clear written documentation habits
• Perfectionist mindset with creativity in finding what others missed and ability to work independently on ambiguous problems
• Ability to engage reliably for ~35 hours per week

Preferred
• Past experience in AI training, model evaluation, or quality review of AI-generated work

How to Apply
Click "Apply" to be taken to the Mercor website. Complete your profile by uploading your resume and confirming your work location. Once verified, you will be matched to opportunities as they arise. Please note that this role cannot support H1B or STEM OPT candidates. Applying through our link supports WFH Bulletin as a referral partner, but you are welcome to apply directly if you prefer.

Responsibilities
• Design test cases that confirm each task works as intended, including tricky edge cases
• Review tasks and reference solutions before finalization, catching ambiguity and gaps early
• Debug tasks or checks in Python when they don’t behave as expected
• Help build simple, repeatable quality checklists and provide actionable feedback to authors
• Watch for shortcuts and grading gaps in AI agent runs to keep benchmark scores trustworthy

Requirements
• MSc or PhD in a STEM field, or equivalent practical experience in a research-heavy or engineering-heavy domain
• 1+ years of experience in test engineering, quality assurance, or a research/software engineering role with strong quality ownership
• Demonstrated skill designing test cases and quality-review processes, and debugging complex systems end-to-end
• Working proficiency in Python and Git, with comfort navigating unfamiliar codebases and environments
• Exceptional attention to detail and clear written documentation habits
• Perfectionist mindset with creativity in finding what others missed and ability to work independently on ambiguous problems
• Ability to engage reliably for ~35 hours per week

Preferred
• Past experience in AI training, model evaluation, or quality review of AI-generated work

How to Apply
Click "Apply" to be taken to the Mercor website. Complete your profile by uploading your resume and confirming your work location. Once verified, you will be matched to opportunities as they arise. Please note that this role cannot support H1B or STEM OPT candidates. Applying through our link supports WFH Bulletin as a referral partner, but you are welcome to apply directly if you prefer.

Responsibilities
• Design test cases that confirm each task works as intended, including tricky edge cases
• Review tasks and reference solutions before finalization, catching ambiguity and gaps early
• Debug tasks or checks in Python when they don’t behave as expected
• Help build simple, repeatable quality checklists and provide actionable feedback to authors
• Watch for shortcuts and grading gaps in AI agent runs to keep benchmark scores trustworthy

Requirements
• MSc or PhD in a STEM field, or equivalent practical experience in a research-heavy or engineering-heavy domain
• 1+ years of experience in test engineering, quality assurance, or a research/software engineering role with strong quality ownership
• Demonstrated skill designing test cases and quality-review processes, and debugging complex systems end-to-end
• Working proficiency in Python and Git, with comfort navigating unfamiliar codebases and environments
• Exceptional attention to detail and clear written documentation habits
• Perfectionist mindset with creativity in finding what others missed and ability to work independently on ambiguous problems
• Ability to engage reliably for ~35 hours per week

Preferred
• Past experience in AI training, model evaluation, or quality review of AI-generated work

How to Apply
Click "Apply" to be taken to the Mercor website. Complete your profile by uploading your resume and confirming your work location. Once verified, you will be matched to opportunities as they arise. Please note that this role cannot support H1B or STEM OPT candidates. Applying through our link supports WFH Bulletin as a referral partner, but you are welcome to apply directly if you prefer.

Take the next step

Take the next step

Mercor

QA/Test Engineer

QA/Test Engineer

Overview

Overview

Cincinnatus LLC is recruiting QA/Test Engineers for a leading AI lab to act as the quality backbone for complex multi-step benchmark tasks. You will design test cases, review tasks and reference solutions, debug issues, and help build repeatable quality processes. This is a full-time W-2 contingent role (~35 hours/week) with the opportunity to be placed at a leading AI lab.

Cincinnatus LLC is recruiting QA/Test Engineers for a leading AI lab to act as the quality backbone for complex multi-step benchmark tasks. You will design test cases, review tasks and reference solutions, debug issues, and help build repeatable quality processes. This is a full-time W-2 contingent role (~35 hours/week) with the opportunity to be placed at a leading AI lab.

Cincinnatus LLC is recruiting QA/Test Engineers for a leading AI lab to act as the quality backbone for complex multi-step benchmark tasks. You will design test cases, review tasks and reference solutions, debug issues, and help build repeatable quality processes. This is a full-time W-2 contingent role (~35 hours/week) with the opportunity to be placed at a leading AI lab.

Get Started

Find Verified Remote Jobs That Fit Your Career Goals

Explore carefully reviewed remote job opportunities from trusted companies worldwide. Discover roles that match your skills, experience and work preferences all in one place.

CAREERS

Find Remote Roles That Match Your Skills

BLOG

Insights, Tips, and Trends for Remote Careers

Newsletter

Get Started

Find Verified Remote Jobs That Fit Your Career Goals

Explore carefully reviewed remote job opportunities from trusted companies worldwide. Discover roles that match your skills, experience and work preferences all in one place.

CAREERS

Find Remote Roles That Match Your Skills

BLOG

Insights, Tips, and Trends for Remote Careers

Newsletter

Get Started

Find Verified Remote Jobs That Fit Your Career Goals

Explore carefully reviewed remote job opportunities from trusted companies worldwide. Discover roles that match your skills, experience and work preferences all in one place.

CAREERS

Find Remote Roles That Match Your Skills

BLOG

Insights, Tips, and Trends for Remote Careers

Newsletter

© 2026 WFH Bulletin. All rights reserved.

© 2026 WFH Bulletin. All rights reserved.