Retour à la recherche
H
HiredSource d’offres vérifiée

Python Engineer (Remote)

Offre en anglais

Design and maintain high-quality Python code to train, optimize, and benchmark large language models. Lead supervised fine-tuning efforts and collaborate on reinforcement learning with human feedback to refine reward models.

  • Télétravail
  • Canada
  • Publié 25 août 2026
  • 1 poste

D’autres postes auxquels postuler directement

Des possibilités semblables publiées par des employeurs qui recrutent sur Jobs.ca, sans formulaire externe.

Résumé du poste

Role: Python Engineer (Remote) Location: Remote (Work from Anywhere) Job Type: Full-Time Payout: Competitive, based on experience Role Overview: We are hiring for one of our clients, seeking a Senior Python Developer to assist a foundational LLM company in enhancing their large language models. The goal is to provide high-quality proprietary data for fine-tuning and benchmarking model performance. Key Responsibilities: • Design, develop, and maintain efficient, high-quality Python code to train and optimize AI models. • Conduct evaluations to benchmark model performance and analyze results for continuous improvement. • Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria. • Lead efforts in supervised fine-tuning, including creating and maintaining high-quality, task-specific datasets. • Collaborate with researchers and annotators to execute reinforcement learning with human feedback and refine reward models. Required Skills & Qualifications: • Proficiency in Python and related frameworks/libraries for AI/ML tasks. • Experience with supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF). • Strong understanding of evaluation strategies and benchmarking processes for AI models. • Ability to design and implement Python code for data generation and model optimization. • Familiarity with AI model response evaluation and ranking methodologies. More About the Opportunity: This role offers a unique opportunity to work with a global leader in artificial intelligence, contributing to the advancement of large language models. Candidates will collaborate with top researchers and engineers in the field. Equal Opportunity Employer: We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications. Apply Now!

Ce que vous ferez

Design and maintain high-quality Python code to train, optimize, and benchmark large language models. Lead supervised fine-tuning efforts and collaborate on reinforcement learning with human feedback to refine reward models.

Exigences

Requires proficiency in Python and experience with SFT and RLHF methodologies. Candidates must have a strong understanding of AI model evaluation strategies and benchmarking processes.

Compétences indiquées

  • PythonSouhaitée

Autres compétences pertinentes

Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.

  • Python
  • Supervised Fine-Tuning
  • RLHF
  • AI Model Benchmarking
  • Data Generation
  • Model Optimization
  • Large Language Models
  • AI Evaluation

Domaines d’emploi

  • Software
  • Technology
  • Engineering
  • Science & Research
  • Data & Analytics

Renseignements supplémentaires

Expérience minimale
5+ ans
Langue de l’offre
anglais
Heures de travail
40 heures par semaine
Niveau d’expérience
Associate