Copy of LLM Model Response Evaluation
lifted an upwork companytmDenver, CO
Copy of LLM Model Response Evaluation
lifted an upwork companytmDenver, CO
yesterday
Industries
Other Scientific and Technical Consulting ServicesAll Other Professional, Scientific, and Technical ServicesCustom Computer Programming ServicesAbout the role
Copy Of LLM Model Response Evaluation Our client, a global technology company that helps businesses build, train, and manage AI systems is looking for experts to evaluate model-generated content against defined quality rubrics such as factuality, consistency, aesthetics, and other evaluation criteria. Evaluating UI widgets, infographics, image factuality, side-by-side comparisons, and similar AI evaluation activities. The work may involve text, images, audio, video, HTML widgets, PDFs, or combinations of these modalities. The work is domain-agnostic and may cover topics across arts, culture, history, science, engineering, and more. Resources will be expected to independently research unfamiliar topics using trusted sources before making evaluation decisions. Each task will include detailed project guidelines within the evaluation platform.3+ years of hands-on experience in LLM / GenAI data evaluation. Bachelor's Degree required. Ability to research unfamiliar topics using trusted sources and make well-supported judgments. Comfortable evaluating content across multiple modalities. Flexible and remote work. Variable workload: Accept or decline tasks based on your availability. No guaranteed hours: Workload may vary weekly.
Matching similar jobs
JOB OVERVIEW
Experience level
Mid
Location
Denver, CO
Occupation
Writers and Authors
Industry
Other Scientific and Technical Consulting Services
Posted
yesterday
Tired of running searches?
Rank the roles you'd take once, and matches like these arrive on their own.
CREATE PROFILE