Lifted (an Upwork Company)

LLM Model Response Evaluation

United States · REMOTE · FREELANCE
Publiée le 1 octobre 2026 · Candidature traitée sur le site de l’entreprise
Data ScienceAI-EvaluatorLLM-EvaluatorAI-TrainerContent-EvaluatorAI-Annotation-SpecialistAI-Response-EvaluationLLM-EvaluationAI-LLM-EvaluationLanguage-Model-EvaluationAI-Language-Model-EvaluationAI-model-evaluationAI-ML-Model-EvaluationAI-Response-EvaluatorAI-Model-Assessment

• Evaluating UI widgets, infographics, image factuality, side-by-side comparisons, and similar AI evaluation activities. • The work may involve text, images, audio, video, HTML widgets, PDFs, or combinations of these modalities. • The work is domain-agnostic and may cover topics across arts, culture, history, science, engineering, and more. • Resources will be expected to independently research unfamiliar topics using trusted sources before making evaluation decisions. • Each task will include detailed project guidelines within the evaluation platform. • 3+ years of hands-on experience in LLM / GenAI data evaluation. • Bachelor's Degree required • Ability to research unfamiliar topics using trusted sources and make well-supported judgments. • Comfortable evaluating content across multiple modalities Flexible and remote work Variable workload: Accept or decline tasks based on your availability No guaranteed hours: Workload may vary weekly Our client, a global technology company that helps businesses build, train, and manage AI systems is looking for experts to evaluate model-generated content against defined quality rubrics such as factuality, consistency, aesthetics, and other evaluation criteria. Originally posted on Himalayas