About the role
QA Engineer – AI Training & Evaluation
Experience: 3–8 Years
Job Type: Contractor
Location: Remote
Compensation: $30–$60/hour
Role Overview:
Contribute to LLM training and evaluation workflows by testing software, reviewing AI-generated outputs, and validating model behavior. Assess outputs for functional correctness, accuracy, reliability, edge-case handling, and adherence to requirements. Create and review test cases, test scenarios, evaluation criteria, bug reports, and quality benchmarks for AI-driven workflows. Provide detailed feedback to improve LLM performance, software quality, and reliability across complex technical tasks.
Key Requirements:
3–8 years of professional experience in QA, Software Testing, Quality Engineering, or related roles. Strong experience with manual and automated testing, test case design, debugging, API testing, and regression testing. Proficiency with Python, Java Script, Java, or similar languages, along with tools such as Selenium, Playwright, Cypress, or equivalent. Experience with or strong interest in LLM evaluation, AI-assisted testing, prompt engineering, or AI training workflows.
Matching similar jobs
JOB OVERVIEW
Salary
$30–$60/hour
Experience level
Senior
Location
Denver, CO
Occupation
Software Quality Assurance Analysts and Testers
Industry
Custom Computer Programming Services
Posted
2 days ago
Tired of running searches?
Rank the roles you'd take once, and matches like these arrive on their own.
CREATE PROFILE