About the role

QA Engineer – AI Training & Evaluation
Experience: 3–8 Years
Job Type: Contractor
Location: Remote
Compensation: $30–$60/hour
Role Overview: Contribute to LLM training and evaluation workflows by testing software, reviewing AI-generated outputs, and validating model behavior. Assess outputs for functional correctness, accuracy, reliability, edge-case handling, and adherence to requirements. Create and review test cases, test scenarios, evaluation criteria, bug reports, and quality benchmarks for AI-driven workflows. Provide detailed feedback to improve LLM performance, software quality, and reliability across complex technical tasks.
Key Requirements: 3–8 years of professional experience in QA, Software Testing, Quality Engineering, or related roles. Strong experience with manual and automated testing, test case design, debugging, API testing, and regression testing. Proficiency with Python, Java Script, Java, or similar languages, along with tools such as Selenium, Playwright, Cypress, or equivalent. Experience with or strong interest in LLM evaluation, AI-assisted testing, prompt engineering, or AI training workflows.

Matching similar jobs

JOB OVERVIEW

Salary

$30–$60/hour

Experience level

Senior

Location

Denver, CO

Occupation

Software Quality Assurance Analysts and Testers

Industry

Custom Computer Programming Services

Posted

2 days ago

Tired of running searches?

Rank the roles you'd take once, and matches like these arrive on their own.

CREATE PROFILE