Copy of LLM Model Response Evaluation
Lifted (an Upwork Company) · Ottawa
Publié le 6 août 2026
Vous serez redirigé vers le site d'origine de l'offre.
Description du poste
About the Role
Evaluating UI widgets, infographics, image factuality, side-by-side comparisons, and similar AI evaluation activities. The work may involve text, images, audio, video, HTML widgets, PDFs, or combinations of these modalities. The work is domain-agnostic and may cover topics across arts, culture, history, science, engineering, and more. Resources will be expected to independently research unfamiliar topics using trusted sources before making evaluation decisions. Each task will include detailed project guidelines within the evaluation platform.
Qualifications
- 3+ years of hands-on experience in LLM / GenAI data evaluation.
- Bachelor's Degree required
- Ability to research unfamiliar topics using trusted sources and make well-supported judgments.
- Comfortable evaluating content across multiple modalities
Conditions
- Flexible and remote work
- Variable workload: Accept or decline tasks based on your availability
- No guaranteed hours: Workload may vary weekly
Mise en page générée automatiquement à partir du texte de l'annonce. En cas de doute, l'offre d'origine fait foi.
Autres offres chez Lifted (an Upwork Company)
Lifted (an Upwork Company) recrute aussi pour un autre poste dans la région :
- Copy of Senior Python Developer (AI Evaluation & Benchmarking) · Région de la capitale nationale