Copy of LLM Model Response Evaluation
Offre en anglaisEvaluate AI-generated content across various modalities including text, images, and video against quality rubrics for factuality and consistency. Conduct independent research on unfamiliar topics to ensure accurate and well-supported evaluation decisions.
- Télétravail
- Ottawa, Ontario, Canada
- Publié 6 août 2026
- Postuler avant le 2 févr. 2027
- 1 poste
D’autres postes auxquels postuler directement
Des possibilités semblables publiées par des employeurs qui recrutent sur Jobs.ca, sans formulaire externe.
Résumé du poste
Our client, a global technology company that helps businesses build, train, and manage AI systems is looking for experts to evaluate model-generated content against defined quality rubrics such as factuality, consistency, aesthetics, and other evaluation criteria. Job Description Evaluating UI widgets, infographics, image factuality, side-by-side comparisons, and similar AI evaluation activities. The work may involve text, images, audio, video, HTML widgets, PDFs, or combinations of these modalities. The work is domain-agnostic and may cover topics across arts, culture, history, science, engineering, and more. Resources will be expected to independently research unfamiliar topics using trusted sources before making evaluation decisions. Each task will include detailed project guidelines within the evaluation platform. Qualifications 3+ years of hands-on experience in LLM / GenAI data evaluation. Bachelor's Degree required Ability to research unfamiliar topics using trusted sources and make well-supported judgments. Comfortable evaluating content across multiple modalities Additional Information Flexible and remote work Variable workload: Accept or decline tasks based on your availability No guaranteed hours: Workload may vary weekly
Ce que vous ferez
Evaluate AI-generated content across various modalities including text, images, and video against quality rubrics for factuality and consistency. Conduct independent research on unfamiliar topics to ensure accurate and well-supported evaluation decisions.
Exigences
Requires a Bachelor's degree and at least 3 years of hands-on experience in LLM or Generative AI data evaluation. Candidates must be comfortable working with multiple content modalities and performing independent research.
Compétences indiquées
- Analyse de donnéesSouhaitée
Autres compétences pertinentes
Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.
- LLM Evaluation
- Generative AI Data Evaluation
- Factuality Assessment
- Multimodal Content Evaluation
- Research
- Quality Rubric Application
- Data Analysis
Domaines d’emploi
- Data & Analytics
- Technology
- Software
- Science & Research
- Consulting
Renseignements supplémentaires
- Formation minimale
- Baccalauréat
- Expérience minimale
- 2+ ans
- Postuler avant le
- 2 févr. 2027
- Langue de l’offre
- anglais
- Heures de travail
- 40 heures par semaine
- Exigences de lieu
- Country, Ottawa, Ontario, Canada
- Niveau d’expérience
- Not Applicable
- Mode de candidature
- La candidature directe est offerte