Muruganantham Ganesan
Open to opportunitiesSenior Application Support Engineer | Site Reliability Engineer (SRE)
Toronto, ON
About
Senior Application Support Engineer | Site Reliability Engineer (SRE) with 13+ years of experience supporting mission-critical Banking and Telecom applications across APAC, the US, and Mexico. Expertise in Production Support, P1/P2 Incident Management, Root Cause Analysis (RCA), Splunk, Linux, SQL, Observability, AWS, and DevOps practices. Proven track record of reducing MTTR by 30%, ensuring 99.9% application availability, and delivering reliable 24×7 production support. AWS Certified Solutions Architect with hands-on Cloud & DevOps expertise. Open to relocation and available for immediate joining.
Skills
- Ability to work under pressure
- adaptability
- Agile
- Amazon Web Services
- Attention to detail
- Business Application Support
- Change Management
- CI/CD
- Communication
- Critical Thinking
- Cross-Functional Collaboration
- DNS
- Docker
- Git
- GitHub
- Incident Management
- Java
- Jira
- Kubernetes
- Linux
- MySQL
- Oracle
- Problem solving
- Project management
- REST APIs
- Root Cause Analysis
- Scrum
- Splunk
- Spring Boot
- SQL
- Stakeholder Management
- System monitoring
- TCP/IP
- Team Leadership
- Teamwork
- Technical Documentation
- Terraform
- Troubleshooting
Experience
Cloud & DevOps Professional Development
Self-Directed Learning / CloudTrain
Nov 2025 to Present
Remote
• Completed a comprehensive Cloud & DevOps Engineering program covering AWS, Linux, Docker, Kubernetes, Jenkins, Git, Terraform, Splunk, Prometheus, Grafana, CI/CD, and Infrastructure as Code (IaC). • Completed 10+ hands-on Cloud & DevOps labs covering CI/CD pipelines, AWS infrastructure provisioning, Docker containerization, and Kubernetes deployments. • Designed and developed 2 Splunk Enterprise dashboards and configured Prometheus and Grafana for application monitoring, log analysis, alerting, and observability using simulated banking application environments. • Strengthened practical expertise in Linux Administration, Shell Scripting, SQL Troubleshooting, Production Support, Incident Management, Cloud Operations, and Site Reliability Engineering (SRE) through continuous hands-on lab exercises.
Project Lead | Senior Application Support Engineer (Site Reliability Engineering (SRE) & Production Operations)
HCL Tech Mexico | USAA
Sep 2022 to Sep 2025
Mexico / India
• Led end-to-end P1/P2 Incident Management for mission-critical banking applications, ensuring rapid service restoration, 99.9% SLA compliance, and high application availability across global production environments. • Managed 15+ monthly major incidents by coordinating bridge calls across Infrastructure, Cloud, Middleware, Database, Application, and Vendor teams, accelerating incident resolution and minimizing business impact. • Performed Risk Analysis, Impact Analysis, and Root Cause Analysis (RCA), reducing recurring production incidents by 40% and improving service restoration time by 25% through operational improvements. • Delivered executive-level stakeholder communication, incident timelines, risk assessments, and escalation updates to business and technology leadership throughout the incident lifecycle. • Coordinated 24×7 Follow-the-Sun production support and seamless handovers across APAC, USA, and Mexico, ensuring operational continuity for globally distributed teams. • Partnered with Application Support, SRE, Infrastructure, Change Management, and Problem Management teams to improve production stability, reduce operational risk, and drive continuous service improvements.
Senior Application Support Engineer (Production Support & Operations)
HCL Technologies | USAA
Jan 2017 to Aug 2022
Chennai, India
• Performed JVM troubleshooting and GC tuning, reducing recurring production incidents by 40% across enterprise banking systems. • Optimized Oracle and DB2 SQL queries, improving batch processing reliability and overall application performance and service availability. • Implemented Prometheus and Grafana monitoring dashboards, reducing MTTD by 30% across production environments. • Automated operational support tasks using Shell scripting, improving response efficiency and reducing manual operational effort. • Collaborated with infrastructure and release teams to reduce production deployment failures by 25% and improve operational stability.
Senior Software Developer / Production Support Engineer
HCL Technologies | Staples Inc.
Dec 2013 to Dec 2016
Chennai, India
• Developed Java and REST API components, improving application response time by 20% across enterprise retail systems. • Resolved high-priority production incidents through debugging and log analysis, reducing recurring issues by 25%. • Supported high-volume production systems processing 1,000+ daily transactions with stable application availability and operational continuity.
Senior Software Developer
Infinite Computer Solutions | Verizon Data Services India
Feb 2012 to Dec 2012
Chennai, India
• Developed enterprise application modules supporting 500+ daily transactions across production and testing environments. • Reduced post-release defects by 20% through validation testing, issue analysis, and deployment support activities. • Supported release coordination and production fixes, improving deployment reliability and operational stability by 15%.
Senior Developer
Prodapt Solutions Pvt Ltd | Windstream Communication
Jul 2010 to Apr 2011
Chennai, India
• Supported EFT and card transaction systems with 99% system availability across critical business operations. • Resolved Sev1 and Sev2 incidents within SLA timelines through troubleshooting and coordinated support activities. • Improved operational stability through structured RCA, monitoring improvements, and recurring issue prevention initiatives.
Software Developer
Infonovum Technologies Pvt Ltd | : ewmglobal.com
Aug 2008 to Jun 2010
Chennai, India
• Supported enterprise applications processing 1,000+ daily transactions while maintaining stable production operations. • Resolved 20+ monthly production issues through troubleshooting and log analysis, reducing downtime by 15%. • Supported deployment and enhancement activities across releases, reducing operational processing errors by 18%.
Education
Bharathidasan University
MCA
Tamil Nadu
2003
Licences & certifications
Cloud & DevOps Engineering Program
CloudTrain
AWS Certified Solutions Architect – Associate
AWS
SCJP 1.5
