Retour à la recherche
JG
J&M GroupSource d’offres vérifiée

Senior Cloud Support Engineer

Offre en anglais

Provide L2/L3 production support for enterprise cloud platforms and business-critical applications across multi-cloud environments. Monitor infrastructure using observability tools and perform root cause analysis to ensure service availability and operational excellence.

  • Sur place
  • Toronto, ON
  • Publié 26 août 2026
  • Postuler avant le 22 févr. 2027
  • 1 poste

D’autres postes auxquels postuler directement

Des possibilités semblables publiées par des employeurs qui recrutent sur Jobs.ca, sans formulaire externe.

Résumé du poste

We are hiring for a Senior Cloud Support Engineer. The role focuses on L2/L3 production support for enterprise cloud platforms and business-critical applications across Azure, AWS, and Google Cloud environments. A strong fit will have 5+ years of cloud or production support experience, multi-cloud expertise, monitoring and observability skills, and strong incident management capabilities. What You Bring Strong cloud certifications such as Microsoft Azure, AWS, and/or Google Cloud. 5+ years of experience in Cloud Support, Production Support, or Site Reliability Operations. Hands-on experience with Dynatrace, Zabbix, Azure Monitor, Log Analytics, and enterprise monitoring tools. Experience supporting cloud-hosted applications and infrastructure across Azure, AWS, and Google Cloud. Strong incident management, major incident response, problem management, and root cause analysis skills. Experience supporting Kubernetes-based applications and containerized workloads. Ability to analyze application logs, metrics, alerts, dashboards, and system health indicators. Experience working in 24x7 production environments with SLA-driven support models. Strong troubleshooting skills for APIs, distributed systems, and enterprise SaaS platforms. Experience with ServiceNow, Jira, Confluence, runbooks, and operational documentation. What you'll do Monitor cloud infrastructure, applications, and services using enterprise observability platforms. Provide L2/L3 support and resolve complex production incidents. Perform root cause analysis and coordinate corrective actions with engineering teams. Support Kubernetes environments, application deployments, and service availability activities. Analyze application, infrastructure, and database performance issues. Manage incident communications and ensure timely service restoration. Develop and maintain operational runbooks, SOPs, and knowledge articles. Collaborate with cloud, infrastructure, development, and product teams to improve operational excellence. Nice to have Strong SQL skills including data analysis, database troubleshooting, and production data fixes. Linux administration and troubleshooting experience. Experience with MySQL and PostgreSQL. Experience with LDAP, access management, and infrastructure provisioning. Exposure to Splunk and other observability platforms. Cloud marketplace, SaaS commerce, or enterprise application support experience. Release validation, change management, and deployment support experience.

Ce que vous ferez

Provide L2/L3 production support for enterprise cloud platforms and business-critical applications across multi-cloud environments. Monitor infrastructure using observability tools and perform root cause analysis to ensure service availability and operational excellence.

Exigences

Requires 5+ years of experience in cloud or production support with strong certifications in Azure, AWS, or GCP. Must have hands-on experience with Kubernetes, enterprise monitoring tools, and incident management in 24x7 environments.

Compétences indiquées

  • Microsoft AzureSouhaitée
  • KubernetesSouhaitée
  • JiraSouhaitée
  • Amazon Web ServicesSouhaitée

Autres compétences pertinentes

Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.

  • Azure
  • AWS
  • Google Cloud Platform
  • Production Support
  • Incident Management
  • Kubernetes
  • Dynatrace
  • Zabbix
  • Azure Monitor
  • Log Analytics
  • Root Cause Analysis
  • SRE
  • ServiceNow
  • Jira
  • API Troubleshooting
  • Distributed Systems

Domaines d’emploi

  • Technology
  • Software
  • Engineering
  • Customer Service & Support
  • Consulting

Renseignements supplémentaires

Expérience minimale
5+ ans
Postuler avant le
22 févr. 2027
Langue de l’offre
anglais
Heures de travail
40 heures par semaine
Niveau d’expérience
Mid-Senior level
Mode de candidature
La candidature directe est offerte