Back to job search
JG
J&M GroupVerified Job Source

Senior Cloud Support Engineer

Provide L2/L3 production support for enterprise cloud platforms and business-critical applications across multi-cloud environments. Monitor infrastructure using observability tools and perform root cause analysis to ensure service availability and operational excellence.

  • On-site
  • Toronto, ON
  • Posted Aug 26, 2026
  • Apply by Feb 22, 2027
  • 1 position

More jobs you can apply to directly

Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.

Job summary

We are hiring for a Senior Cloud Support Engineer. The role focuses on L2/L3 production support for enterprise cloud platforms and business-critical applications across Azure, AWS, and Google Cloud environments. A strong fit will have 5+ years of cloud or production support experience, multi-cloud expertise, monitoring and observability skills, and strong incident management capabilities. What You Bring Strong cloud certifications such as Microsoft Azure, AWS, and/or Google Cloud. 5+ years of experience in Cloud Support, Production Support, or Site Reliability Operations. Hands-on experience with Dynatrace, Zabbix, Azure Monitor, Log Analytics, and enterprise monitoring tools. Experience supporting cloud-hosted applications and infrastructure across Azure, AWS, and Google Cloud. Strong incident management, major incident response, problem management, and root cause analysis skills. Experience supporting Kubernetes-based applications and containerized workloads. Ability to analyze application logs, metrics, alerts, dashboards, and system health indicators. Experience working in 24x7 production environments with SLA-driven support models. Strong troubleshooting skills for APIs, distributed systems, and enterprise SaaS platforms. Experience with ServiceNow, Jira, Confluence, runbooks, and operational documentation. What you'll do Monitor cloud infrastructure, applications, and services using enterprise observability platforms. Provide L2/L3 support and resolve complex production incidents. Perform root cause analysis and coordinate corrective actions with engineering teams. Support Kubernetes environments, application deployments, and service availability activities. Analyze application, infrastructure, and database performance issues. Manage incident communications and ensure timely service restoration. Develop and maintain operational runbooks, SOPs, and knowledge articles. Collaborate with cloud, infrastructure, development, and product teams to improve operational excellence. Nice to have Strong SQL skills including data analysis, database troubleshooting, and production data fixes. Linux administration and troubleshooting experience. Experience with MySQL and PostgreSQL. Experience with LDAP, access management, and infrastructure provisioning. Exposure to Splunk and other observability platforms. Cloud marketplace, SaaS commerce, or enterprise application support experience. Release validation, change management, and deployment support experience.

What you’ll do

Provide L2/L3 production support for enterprise cloud platforms and business-critical applications across multi-cloud environments. Monitor infrastructure using observability tools and perform root cause analysis to ensure service availability and operational excellence.

Requirements

Requires 5+ years of experience in cloud or production support with strong certifications in Azure, AWS, or GCP. Must have hands-on experience with Kubernetes, enterprise monitoring tools, and incident management in 24x7 environments.

Listed skills

  • Microsoft AzurePreferred
  • KubernetesPreferred
  • JiraPreferred
  • Amazon Web ServicesPreferred

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • Azure
  • AWS
  • Google Cloud Platform
  • Production Support
  • Incident Management
  • Kubernetes
  • Dynatrace
  • Zabbix
  • Azure Monitor
  • Log Analytics
  • Root Cause Analysis
  • SRE
  • ServiceNow
  • Jira
  • API Troubleshooting
  • Distributed Systems

Job areas

  • Technology
  • Software
  • Engineering
  • Customer Service & Support
  • Consulting

Additional details

Minimum experience
5+ years
Apply by
Feb 22, 2027
Posting language
English
Working hours
40 hours per week
Seniority
Mid-Senior level
Application method
Direct apply is available