MG

Muruganantham Ganesan

Ouvert aux possibilités

Senior Application Support Engineer | Site Reliability Engineer (SRE)

Toronto, ON

Self-Directed Learning / CloudTrain
Bharathidasan University

À propos

Senior Application Support Engineer | Site Reliability Engineer (SRE) with 13+ years of experience supporting mission-critical Banking and Telecom applications across APAC, the US, and Mexico. Expertise in Production Support, P1/P2 Incident Management, Root Cause Analysis (RCA), Splunk, Linux, SQL, Observability, AWS, and DevOps practices. Proven track record of reducing MTTR by 30%, ensuring 99.9% application availability, and delivering reliable 24×7 production support. AWS Certified Solutions Architect with hands-on Cloud & DevOps expertise. Open to relocation and available for immediate joining.

Compétences

  • Ability to work under pressure
  • adaptability
  • Agile
  • Amazon Web Services
  • Souci du détail
  • Business Application Support
  • Change Management
  • CI/CD
  • Communication
  • Critical Thinking
  • Cross-Functional Collaboration
  • DNS
  • Docker
  • Git
  • GitHub
  • Incident Management
  • Java
  • Jira
  • Kubernetes
  • Linux
  • MySQL
  • Oracle
  • Résolution de problèmes
  • Gestion de projet
  • API REST
  • Root Cause Analysis
  • Scrum
  • Splunk
  • Spring Boot
  • SQL
  • Stakeholder Management
  • System monitoring
  • TCP/IP
  • Team Leadership
  • Travail d’équipe
  • Technical Documentation
  • Terraform
  • Troubleshooting

Expérience

  1. Cloud & DevOps Professional Development

    Self-Directed Learning / CloudTrain

    nov. 2025 to Aujourd’hui

    Remote

    • Completed a comprehensive Cloud & DevOps Engineering program covering AWS, Linux, Docker, Kubernetes, Jenkins, Git, Terraform, Splunk, Prometheus, Grafana, CI/CD, and Infrastructure as Code (IaC). • Completed 10+ hands-on Cloud & DevOps labs covering CI/CD pipelines, AWS infrastructure provisioning, Docker containerization, and Kubernetes deployments. • Designed and developed 2 Splunk Enterprise dashboards and configured Prometheus and Grafana for application monitoring, log analysis, alerting, and observability using simulated banking application environments. • Strengthened practical expertise in Linux Administration, Shell Scripting, SQL Troubleshooting, Production Support, Incident Management, Cloud Operations, and Site Reliability Engineering (SRE) through continuous hands-on lab exercises.

  2. Project Lead | Senior Application Support Engineer (Site Reliability Engineering (SRE) & Production Operations)

    HCL Tech Mexico | USAA

    sept. 2022 to sept. 2025

    Mexico / India

    • Led end-to-end P1/P2 Incident Management for mission-critical banking applications, ensuring rapid service restoration, 99.9% SLA compliance, and high application availability across global production environments. • Managed 15+ monthly major incidents by coordinating bridge calls across Infrastructure, Cloud, Middleware, Database, Application, and Vendor teams, accelerating incident resolution and minimizing business impact. • Performed Risk Analysis, Impact Analysis, and Root Cause Analysis (RCA), reducing recurring production incidents by 40% and improving service restoration time by 25% through operational improvements. • Delivered executive-level stakeholder communication, incident timelines, risk assessments, and escalation updates to business and technology leadership throughout the incident lifecycle. • Coordinated 24×7 Follow-the-Sun production support and seamless handovers across APAC, USA, and Mexico, ensuring operational continuity for globally distributed teams. • Partnered with Application Support, SRE, Infrastructure, Change Management, and Problem Management teams to improve production stability, reduce operational risk, and drive continuous service improvements.

  3. Senior Application Support Engineer (Production Support & Operations)

    HCL Technologies | USAA

    janv. 2017 to août 2022

    Chennai, India

    • Performed JVM troubleshooting and GC tuning, reducing recurring production incidents by 40% across enterprise banking systems. • Optimized Oracle and DB2 SQL queries, improving batch processing reliability and overall application performance and service availability. • Implemented Prometheus and Grafana monitoring dashboards, reducing MTTD by 30% across production environments. • Automated operational support tasks using Shell scripting, improving response efficiency and reducing manual operational effort. • Collaborated with infrastructure and release teams to reduce production deployment failures by 25% and improve operational stability.

  4. Senior Software Developer / Production Support Engineer

    HCL Technologies | Staples Inc.

    déc. 2013 to déc. 2016

    Chennai, India

    • Developed Java and REST API components, improving application response time by 20% across enterprise retail systems. • Resolved high-priority production incidents through debugging and log analysis, reducing recurring issues by 25%. • Supported high-volume production systems processing 1,000+ daily transactions with stable application availability and operational continuity.

  5. Senior Software Developer

    Infinite Computer Solutions | Verizon Data Services India

    févr. 2012 to déc. 2012

    Chennai, India

    • Developed enterprise application modules supporting 500+ daily transactions across production and testing environments. • Reduced post-release defects by 20% through validation testing, issue analysis, and deployment support activities. • Supported release coordination and production fixes, improving deployment reliability and operational stability by 15%.

  6. Senior Developer

    Prodapt Solutions Pvt Ltd | Windstream Communication

    juill. 2010 to avr. 2011

    Chennai, India

    • Supported EFT and card transaction systems with 99% system availability across critical business operations. • Resolved Sev1 and Sev2 incidents within SLA timelines through troubleshooting and coordinated support activities. • Improved operational stability through structured RCA, monitoring improvements, and recurring issue prevention initiatives.

  7. Software Developer

    Infonovum Technologies Pvt Ltd | : ewmglobal.com

    août 2008 to juin 2010

    Chennai, India

    • Supported enterprise applications processing 1,000+ daily transactions while maintaining stable production operations. • Resolved 20+ monthly production issues through troubleshooting and log analysis, reducing downtime by 15%. • Supported deployment and enhancement activities across releases, reducing operational processing errors by 18%.

Formation

  1. Bharathidasan University

    MCA

    Tamil Nadu

    2003

Permis et certifications

  • Cloud & DevOps Engineering Program

    CloudTrain

  • AWS Certified Solutions Architect – Associate

    AWS

  • SCJP 1.5