Retour à la recherche
1L
1851 LabsSource d’offres vérifiée

Machine Learning Engineer

Offre en anglais

Manage a GPU fleet serving millions of images and implement new AI models, papers, or codebases rapidly. Optimize model speed, cost, and features while staying current with AI research.

  • Sur place
  • Toronto, ON
  • Publié 5 août 2026
  • Postuler avant le 4 sept. 2026
  • 1 poste

Résumé du poste

Careers Come build GenTube with us. We make it fun to create things. So fun a billion people will want to do it every day. 100M+ creations so far, millions more every month. 100M+creations made 5people, so far ~2sper creation Torontohome base Open roles Software Engineer (AI, Full Stack) Machine Learning Engineer Research Engineer Designer Operations Software Engineer (AI, Full Stack) A product hacker who cares about what users enjoy, not just code. Can build and launch a full consumer feature alone: app, backend, analytics. Plays with new AI models and turns them into simple, fun experiences. Strong record: great academics, an app with real users, or top-company work. Stack: TypeScript, Node.js, React, Convex, diffusion models, low-latency LLMs. Apply → Machine Learning Engineer Owns a GPU fleet serving millions of images a month. Can get a new model, paper, or codebase working within a day. Knows models deeply: changes layers, boosts speed, cuts costs, builds features. Uses image and video models for fun. Follows new AI research closely. Stack: PyTorch, diffusion models, low-latency LLMs, GPU infra, inference optimization. Apply → Research Engineer Makes models 10–100× faster, cheaper, or better. Knows every optimization trick. Invents the missing ones. Writes the paper, builds the system, proves it works. Curious by default. Stack: PyTorch, CUDA, diffusion models, quantization, distillation, inference optimization. Apply → Designer Believes creating should feel like play, not a form. Designs the core create flows and ships them in code. Lives in the data: cuts what fails, doubles down on what works. Tools: Figma, React/CSS, PostHog. Apply → Operations Sits next to the founders and makes the whole company move faster. Hunts for growth. Builds systems that turn what works into a machine. Relentless, resourceful, allergic to "someone should." Apply → What you get Top-of-market salary, full benefits. Meaningful equity. Backed by top consumer investors. You ship it, the world uses it. The hardest problems in consumer AI. The highest standards. Interested to apply? Just email us: careers@gentube.app

Ce que vous ferez

Manage a GPU fleet serving millions of images and implement new AI models, papers, or codebases rapidly. Optimize model speed, cost, and features while staying current with AI research.

Exigences

Deep knowledge of model layers and inference optimization with a passion for image and video models. Proficiency in PyTorch and experience with GPU infrastructure and low-latency LLMs is required.

Avantages

• Full benefits • Meaningful equity

Autres compétences pertinentes

Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.

  • PyTorch
  • Diffusion models
  • Low-latency LLMs
  • GPU infrastructure
  • Inference optimization
  • Model architecture
  • AI research

Domaines d’emploi

  • Software
  • Technology
  • Engineering
  • Data & Analytics
  • Science & Research

Renseignements supplémentaires

Expérience minimale
2+ ans
Postuler avant le
4 sept. 2026
Langue de l’offre
anglais
Heures de travail
40 heures par semaine
Niveau d’expérience
Mid-Senior level