Retour à la recherche
U
UpworkSource d’offres vérifiée

Copy of LLM Model Response Evaluation

Offre en anglais

Evaluate AI-generated content across various modalities including text, images, and video against quality rubrics for factuality and consistency. Conduct independent research on unfamiliar topics to ensure accurate and well-supported evaluation decisions.

  • Télétravail
  • Ottawa, Ontario, Canada
  • Publié 6 août 2026
  • Postuler avant le 2 févr. 2027
  • 1 poste

D’autres postes auxquels postuler directement

Des possibilités semblables publiées par des employeurs qui recrutent sur Jobs.ca, sans formulaire externe.

Résumé du poste

Our client, a global technology company that helps businesses build, train, and manage AI systems is looking for experts to evaluate model-generated content against defined quality rubrics such as factuality, consistency, aesthetics, and other evaluation criteria. Job Description Evaluating UI widgets, infographics, image factuality, side-by-side comparisons, and similar AI evaluation activities. The work may involve text, images, audio, video, HTML widgets, PDFs, or combinations of these modalities. The work is domain-agnostic and may cover topics across arts, culture, history, science, engineering, and more. Resources will be expected to independently research unfamiliar topics using trusted sources before making evaluation decisions. Each task will include detailed project guidelines within the evaluation platform. Qualifications 3+ years of hands-on experience in LLM / GenAI data evaluation. Bachelor's Degree required Ability to research unfamiliar topics using trusted sources and make well-supported judgments. Comfortable evaluating content across multiple modalities Additional Information Flexible and remote work Variable workload: Accept or decline tasks based on your availability No guaranteed hours: Workload may vary weekly

Ce que vous ferez

Evaluate AI-generated content across various modalities including text, images, and video against quality rubrics for factuality and consistency. Conduct independent research on unfamiliar topics to ensure accurate and well-supported evaluation decisions.

Exigences

Requires a Bachelor's degree and at least 3 years of hands-on experience in LLM or Generative AI data evaluation. Candidates must be comfortable working with multiple content modalities and performing independent research.

Compétences indiquées

  • Analyse de donnéesSouhaitée

Autres compétences pertinentes

Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.

  • LLM Evaluation
  • Generative AI Data Evaluation
  • Factuality Assessment
  • Multimodal Content Evaluation
  • Research
  • Quality Rubric Application
  • Data Analysis

Domaines d’emploi

  • Data & Analytics
  • Technology
  • Software
  • Science & Research
  • Consulting

Renseignements supplémentaires

Formation minimale
Baccalauréat
Expérience minimale
2+ ans
Postuler avant le
2 févr. 2027
Langue de l’offre
anglais
Heures de travail
40 heures par semaine
Exigences de lieu
Country, Ottawa, Ontario, Canada
Niveau d’expérience
Not Applicable
Mode de candidature
La candidature directe est offerte