Back to job search
McGraw Hill logo
McGraw HillVerified Job Source

Senior Site Reliability Engineer

  • Canada
  • Remote
  • Posted Apr 9, 2026
  • 1 position

Opens an external site

Sign in to save this job
Employment type
Full-time
Experience level
Senior · 5+ years
Minimum education
Bachelor’s degree
Apply by
Apr 9, 2027
Posting language
English
Working hours
40 hours per week

Job summary

The Senior Site Reliability Engineer will design, deploy, and automate cloud infrastructure while collaborating with product teams in a DevOps model. They are responsible for ensuring system reliability, performance, security, and scalability through infrastructure-as-code and proactive monitoring.

Job details

Overview Make an Impact! At McGraw Hill we create best-in-class, next-generation learning platforms that are used by millions of students and educators worldwide from kindergarten through graduate school. Our goal is to accelerate student success through intuitive and effective learning tools and content that maximize a teacher’s time and a student’s learning experience. We do all of this in a supportive, collaborative environment where you can grow your career in a way that fits into your life. How can you make an impact? We are hiring a Senior Site Reliability Engineer who will build and support reliable, high-capacity, and well-performing systems in support of our mission to protect and improve the McGraw Hill customer platforms, with an ever-watchful eye on reliability, security, performance, cost, and operational excellence. As a Sr Site Reliability Engineer, you will collaborate in a DevOps model with product development teams; designing, deploying, and managing automation tools that increase predictability as well as time to market while reducing cost This is a remote position open to applicants authorized to work for any employer within Canada. What you will be doing: Partner with product teams in a DevOps model to design, deploy, and automate cloud infrastructure (AWS, Terraform) Optimize system reliability, performance, scalability, and cost efficiency Implement and maintain infrastructure-as-code and monitoring-as-code for transparency and repeatability Own application reliability, uptime, security, capacity, and SLA performance Lead major incident response, on-call support, and triage bridges Enhance observability and telemetry to monitor customer experience, KPIs, and infrastructure health Support secure, agile development practices in partnership with CyberSecurity (DevSecOps) Drive resiliency efforts, including failure testing, capacity forecasting, and scaling plans Mentor engineers and collaborate cross-functionally across stakeholder groups Promote knowledge sharing, automation-first practices, and continuous improvement (Kubernetes/EKS experience preferred) Our Cloud Stack Cloud Platforms: AWS (EC2, ECS, Lambda, S3, VPC, etc.), Azure (VNETs, NSGs, SQL, Monitor), OCI (plus) Infrastructure & OS: Windows Server (2016–2022), AD/Azure AD (AAD Connect), DNS, DHCP, OS hardening & patching Infrastructure as Code & Automation: Terraform, Ansible, Packer, PowerShell Programming & Containers: Python, Golang, Bash; ECS, EKS, OKE Security & Monitoring: Rapid7, WAF, CloudWatch, New Relic, Datadog DevSecOps & Web: Jenkins, CircleCI, Artifactory, SonarQube, GitHub Enterprise; Apache, Tomcat, Angular We are looking for someone with… Experience developing, debugging, and deploying enterprise applications Infrastructure automation and container orchestration experience (e.g., Terraform, EKS, ECS) Strong troubleshooting across web, application, networking, OS, and database technologies Experience with CI/CD, high-concurrency systems, and cloud-based production infrastructure Proven problem-solving, communication, and root cause analysis skills BS in Computer Science or related field (or equivalent experience) Why work for us? The work you do at McGraw Hill will be work that matters. We are collectively designing content that will build the future of education. Play your part and experience a sense of fulfilment that will inspire you to even greater heights. The pay range for this position is between $140,000 - $155,000 CAD annually. However, base pay offered may vary depending on job-related knowledge, skills, experience, and location. An annual bonus plan may be provided as part of the compensation package, in addition to a full range of medical and/or other benefits, depending on the position offered. McGraw Hill recruiters always use a “@mheducation.com” or "@careers.mheducation.com” email addresses and/or from our Applicant Tracking System, iCIMS. Any variation of this email domain should be considered suspicious. Additionally, McGraw Hill recruiters and authorized representatives will never request sensitive information in email. CAN_TECH_26

What you’ll do

The Senior Site Reliability Engineer will design, deploy, and automate cloud infrastructure while collaborating with product teams in a DevOps model. They are responsible for ensuring system reliability, performance, security, and scalability through infrastructure-as-code and proactive monitoring.

Requirements

Candidates must have experience developing, debugging, and deploying enterprise applications with strong skills in infrastructure automation and container orchestration. A Bachelor's degree in Computer Science or equivalent experience is required, along with proficiency in cloud-based production environments.

Benefits

• Medical benefits • Annual bonus plan

Listed skills

  • Microsoft Azure · Preferred
  • Production · Preferred
  • analysis · Preferred
  • Reliability · Preferred
  • Teams · Preferred
  • Development · Preferred
  • Agile · Preferred
  • Amazon Web Services · Preferred
  • Time · Preferred
  • Terraform · Preferred
  • Customer · Preferred
  • email · Preferred

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • AWS
  • Terraform
  • Kubernetes
  • EKS
  • Python
  • Golang
  • Bash
  • Ansible
  • CI/CD
  • DevOps
  • Infrastructure-as-code
  • Observability
  • Troubleshooting
  • Cloud infrastructure
  • System reliability
  • Security

Job areas

  • Technology
  • Software
  • Engineering
  • Education

More jobs you can apply to directly

Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.

Browse all Easy Apply jobs