Software Engineer II
Offre en anglaisDesign and build a secure, vendor-agnostic agent platform including orchestration, memory, and runtime systems. Implement evaluation tooling and safety guardrails to improve agent performance and reliability in production.
- Sur place
- Toronto, ON
- Publié 18 août 2026
- Postuler avant le 17 sept. 2026
- 1 poste
D’autres postes auxquels postuler directement
Des possibilités semblables publiées par des employeurs qui recrutent sur Jobs.ca, sans formulaire externe.
Forgeahead Solutions Corporation
Technical Lead and Senior Software Engineer
- Sur place
BC Public Schools
Manager, People, Performance and Culture
- Sur place
Alcohol and Gaming Commission of Ontario (AGCO)
Responsable de la gestion de l’information
- Sur place
Résumé du poste
Why Socure? Socure is building the identity trust infrastructure for the digital economy — verifying 100% of good identities in real time and stopping fraud before it starts. The mission is big, the problems are complex, and the impact is felt by businesses, governments, and millions of people every day. We hire people who want that level of responsibility. People who move fast, think critically, act like owners, and care deeply about solving customer problems with precision. If you want predictability or narrow scope, this won’t be your place. If you want to help build the future of identity with a team that holds a high bar for itself — keep reading. About the Role The Agentic AI Foundations team is building the core platform, systems, and primitives that enable Socure to transition from traditional software workflows to agent-native operations. As a Software Engineer II on the team, you will help design, build, and harden a secure, evaluable, vendor-agnostic agent platform that teams across Socure can build on, working alongside senior and staff engineers who set the architectural direction. This is a hands-on, zero-to-one team, and you’ll get outsized exposure to how agentic systems are architected and operated in production. You’ll bring strong foundational knowledge of LLMs, agentic AI, and GPU/model serving through academic, research, professional, open-source, or other relevant experience, and grow into greater ownership as you build alongside senior engineers on the team. What You’ll Do Build components of a vendor-agnostic agent platform — including orchestration, tool use, memory, and runtime systems — under the guidance of senior engineers on the team. Implement evaluation and reliability tooling, including metrics, harnesses, and pipelines, to measure and improve agent performance, robustness, and safety in production. Help implement safety and governance controls, including guardrails, policy enforcement, and human-in-the-loop review mechanisms. Build data grounding, retrieval, and memory components that keep agents accurate, context-aware, and aligned with Socure’s domain knowledge and policies. Prototype and iterate on agent behaviors, including planning, multi-step execution, and coordination of tools and services, using real internal workflows as proving grounds. Partner with product and engineering teams to implement agent-powered workflows using the platform primitives the team builds. Apply and help refine documented best practices and design patterns for secure, observable, and scalable agent systems. Bring strong foundational knowledge of LLMs, GPU computing, and model serving to technical discussions and implementation decisions. What You’ll Bring Bachelor’s or Master’s degree in Computer Science, Computer Engineering, Machine Learning/AI, or a related field from top tier institutions, or equivalent practical experience demonstrating strong foundations in computer science and machine learning. 2+ years of professional software engineering experience, with demonstrated experience in distributed systems, backend platforms, infrastructure, or comparable technical environments. Very strong foundational knowledge of large language models and agentic AI systems, including architectures, prompting and orchestration patterns, tool use, and evaluation approaches. Strong foundational understanding of GPU computing and model-serving infrastructure, such as CUDA, vLLM, Ollama, LLMLite, TensorRT-LLM, Triton Inference Server, or similar technologies, including the performance and cost trade-offs associated with serving LLMs at scale. Solid grounding in distributed systems fundamentals, including concurrency, fault tolerance, observability, and performance. Proficiency in at least one modern backend programming language and ecosystem, such as Java, Go, Python, or similar, with comfort working with cloud-native infrastructure, APIs, and data services. Ability to work productively in ambiguous, early-stage problem spaces with guidance from senior engineers, translating direction into working software. A track record of strong technical performance demonstrated through professional impact, research, challenging technical projects, open-source contributions, internships, or other relevant work. Strong collaboration and communication skills, with comfort working alongside cross-functional partners such as product, data science, platform, and security. Preferred Qualifications Experience with multi-agent systems, workflow orchestration, or distributed coordination frameworks through professional work, research, coursework, or technical projects. Experience building or using agent platforms — such as orchestration frameworks, tool registries, or memory systems — or LLM routing, caching, or fine-tuning pipelines through professional work, research, internships, open-source contributions, or personal projects. Exposure to evaluation frameworks, experimentation platforms, or ML systems, such as offline/online evaluations, A/B testing, or agent and model benchmarking. Experience with AI safety, security, or policy systems — including guardrails, policy engines, content filters, or responsible AI frameworks — through professional work, research, coursework, or technical projects. Experience with retrieval systems, knowledge graphs, or data platforms used to ground LLMs and agents in enterprise contexts. Demonstrated depth in ML systems or LLM infrastructure through professional impact, research, publications, technical projects, competition results, open-source contributions, or comparable experience. Socure is an equal opportunity employer that values diversity in all its forms within our company. We do not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. If you need an accommodation during any stage of the application or hiring process—including interview or onboarding support—please reach out to your Socure recruiting partner directly.
Ce que vous ferez
Design and build a secure, vendor-agnostic agent platform including orchestration, memory, and runtime systems. Implement evaluation tooling and safety guardrails to improve agent performance and reliability in production.
Exigences
Requires a degree in Computer Science or AI and 2+ years of professional software engineering experience in distributed systems or backend platforms. Must have strong foundational knowledge of LLMs, GPU computing, and model-serving infrastructure.
Compétences indiquées
- GoSouhaitée
- JavaSouhaitée
- PythonSouhaitée
Autres compétences pertinentes
Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.
- Large Language Models
- Agentic AI
- Distributed Systems
- GPU Computing
- Model Serving
- Python
- Java
- Go
- CUDA
- vLLM
- Triton Inference Server
- Orchestration
- Retrieval Systems
- AI Safety
- Backend Engineering
- Cloud-Native Infrastructure
Domaines d’emploi
- Software
- Technology
- Engineering
- Data & Analytics
- Science & Research
Renseignements supplémentaires
- Formation minimale
- Baccalauréat
- Expérience minimale
- 2+ ans
- Postuler avant le
- 17 sept. 2026
- Langue de l’offre
- anglais
- Heures de travail
- 40 heures par semaine
- Niveau d’expérience
- Mid-Senior level
- Mode de candidature
- La candidature directe est offerte