Retour à la recherche
K
KakeSource d’offres vérifiée

Senior AI/LLM Engineer (Python) - Remote

Offre en anglais

Design, build, and deploy LLM-powered features and backend services using Python and FastAPI. Develop RAG pipelines and orchestration flows while managing model performance, cost, and latency.

  • Télétravail
  • Toronto, Ontario, Canada
  • Publié 13 août 2026
  • Postuler avant le 12 sept. 2026
  • 1 poste

D’autres postes auxquels postuler directement

Des possibilités semblables publiées par des employeurs qui recrutent sur Jobs.ca, sans formulaire externe.

Résumé du poste

Senior AI/LLM Engineer (Python) - Remote Kake is a remote-first company and a people-first global community of senior engineers. Kake engineers are behind some of the world’s most innovative products (Brands you’ve heard!), from Fortune 500 to fast-growing companies. We believe it’s not where your table is, but what you bring to the table that matters. Our community spans 45,000+ engineers across 55+ countries; join a culture where great people stay, grow, and thrive (and love eating kake!). Senior AI/LLM Engineers with strong Python experience, skilled in designing, building, and productionizing LLM-powered applications and AI systems at scale. It's a great fit for people who enjoy working at the intersection of software engineering and applied AI, and who take ownership from prototyping through production deployment. What you’ll build and own Design, build, and deploy LLM-powered features and applications using Python. Develop and maintain backend services and APIs (e.g., FastAPI) that expose AI/LLM capabilities to other systems. Build and optimize RAG pipelines, including embeddings, vector search, and retrieval strategies. Design, test, and iterate on prompts, agents, and orchestration flows using frameworks such as LangChain, LlamaIndex, or similar. Integrate with LLM providers and APIs (e.g., OpenAI, Anthropic, open-source models) and manage tradeoffs around cost, latency, and quality. Evaluate model outputs systematically, building tooling and metrics to test accuracy, safety, and regression across iterations. Work with containerized environments and data infrastructure (e.g., PostgreSQL, Redis, vector databases) to support reliable AI systems in production. Collaborate with cross-functional stakeholders to translate ambiguous product needs into technically sound AI solutions. Core Requirements Strong proficiency in Python and experience with FastAPI or similar backend frameworks. Experience working with LLM APIs (e.g., OpenAI, Anthropic, or similar) and frameworks such as LangChain or LlamaIndex. Experience with RAG architectures, embeddings, and vector databases (e.g., Pinecone, Weaviate, pgvector, or similar). Hands-on experience with Docker and containerized development environments. Experience working with PostgreSQL, Redis, or similar data stores. Strong experience writing functional and integration tests, including evaluation frameworks for AI/LLM output quality. Excellent written and verbal communication skills in English. Ability to work independently in a remote, fast-paced environment. Nice-to-Have Experience fine-tuning or evaluating open-source LLMs. Familiarity with prompt engineering best practices and agentic workflows. Experience with distributed systems, streaming (e.g., Kafka), or large-scale applications. Background in machine learning fundamentals (e.g., scikit-learn, PyTorch, or TensorFlow). Comfortable working flexible hours to overlap with distributed teams across different time zones. Why Join Kake? The icing on the Kake: 💰 Competitive Pay in USD: Work globally, get paid globally. 🌎 Fully Remote: Simply put, we trust you. 💜 Better Me Fund: We invest in your personal growth and passions. 🎂 Compassion is Badass: Join a community that invests in social good. Ready for your piece of the Kake? Apply now! Quick note: Due to the high volume of applications, only shortlisted candidates will be contacted.

Ce que vous ferez

Design, build, and deploy LLM-powered features and backend services using Python and FastAPI. Develop RAG pipelines and orchestration flows while managing model performance, cost, and latency.

Exigences

Requires strong proficiency in Python, experience with LLM frameworks like LangChain, and expertise in vector databases and RAG architectures. Candidates must be skilled in Docker, PostgreSQL, and writing comprehensive evaluation frameworks for AI outputs.

Avantages

• Competitive Pay in USD • Fully Remote • Better Me Fund • Social Good Investment

Compétences indiquées

  • RedisSouhaitée
  • PostgreSQLSouhaitée
  • DockerSouhaitée
  • PythonSouhaitée

Autres compétences pertinentes

Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.

  • Python
  • FastAPI
  • LLM APIs
  • LangChain
  • LlamaIndex
  • RAG
  • Vector Databases
  • Docker
  • PostgreSQL
  • Redis
  • Prompt Engineering
  • Integration Testing
  • AI Evaluation
  • Backend Development
  • API Design
  • Containerization

Domaines d’emploi

  • Software
  • Technology
  • Engineering
  • Data & Analytics
  • Consulting

Renseignements supplémentaires

Expérience minimale
5+ ans
Postuler avant le
12 sept. 2026
Langue de l’offre
anglais
Heures de travail
40 heures par semaine
Exigences de lieu
Country, Toronto, Ontario, Canada
Niveau d’expérience
Mid-Senior level
Mode de candidature
La candidature directe est offerte