Back to job search
Morningstar logo

Senior Site Reliability Engineer

  • Toronto, ON
  • Hybrid
  • Posted Jul 6, 2026
  • 1 position

$90,489–$132,711 / year

Opens an external site

Sign in to save this job
Employment type
Full-time
Experience level
Senior · 5+ years
Minimum education
Bachelor’s degree
Posting language
English
Working hours
40 hours per week
Office presence
4 days per week

Job summary

Design and maintain AWS cloud infrastructure and CI/CD pipelines to ensure platform stability and security. Lead reliability initiatives including disaster recovery, monitoring automation, and collaboration with engineering teams to embed SRE best practices.

Job details

About the Team Investment Services is Morningstar’s internal product group focused on building and maintaining the platforms that power our global data operations. We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third Party Data, and Manager Research teams to collect, process, and deliver high-quality investment data at scale—supporting over 770,000 investments across thousands of global processes. We design and maintain the internal tools and systems that support data collection for operational, performance, portfolio, and document data; enable automation and AI-assisted workflows; improve analyst productivity and experience; and ensure data quality, scalability, and system stability. Our platforms are used by hundreds of analysts across the globe to process billions of data points every month. Location: Toronto, ON (4 days onsite) What You’ll Do Design, build, and improve CI/CD pipelines to accelerate software delivery while maintaining stability and security across our platform. Provision, configure, and maintain cloud infrastructure on AWS using Infrastructure as Code tools such as Terraform, CDK, or CloudFormation. Provide on-call technical triage and troubleshooting, driving incidents to resolution and conducting thorough post-incident reviews. Lead cross-team reliability initiatives, including disaster recovery planning, security compliance, and AWS resource optimization. Deploy and manage containerized applications using Docker and AWS ECS/EKS, optimizing resource utilization and deployment strategies. Drive automation and innovation for proactive monitoring, alerting, and continuous operational improvement using tools such as Splunk, CloudWatch, New Relic, and Harness. Collaborate with software engineers and data engineers to embed SRE best practices into the development lifecycle, including SLOs, error budgets, and capacity planning. Write scripts and tooling in Python, Bash, or other scripting languages to automate routine operational tasks and streamline deployments. Document infrastructure architecture, deployment processes, and operational runbooks to enable transparency, consistency, and long-term maintainability. Collaborate with globally distributed teams for projects, knowledge transfer, and on-call rotation coverage. Leverage AI-assisted development tools (e.g., GitHub Copilot, Claude Code) to accelerate engineering workflows and improve productivity. What We’re Looking For 5+ years of experience in Site Reliability Engineering, DevOps, or cloud infrastructure roles supporting production systems. Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience. Strong hands-on experience with AWS cloud services (EC2, S3, ECS/EKS, Lambda, RDS, VPC, IAM, Route 53, CloudWatch). Proficiency with Infrastructure as Code tools such as Terraform, CDK, or CloudFormation. Experience building and maintaining CI/CD pipelines using tools such as Jenkins, Harness, GitHub Actions, or similar platforms. Strong working knowledge of Docker containers and container orchestration platforms. Proficiency in scripting languages such as Python or Bash for automation and operational tooling. Solid understanding of Linux/Unix system administration and networking fundamentals. Experience with monitoring, logging, and alerting tools such as Splunk, New Relic, CloudWatch, or Datadog. Knowledge of SRE principles, including SLIs/SLOs, error budgets, incident management, and post-incident review processes. Experience using AI-assisted development tools (e.g., GitHub Copilot, Claude Code). Excellent communication and collaboration skills, with the ability to work effectively across distributed teams and explain infrastructure decisions clearly Nice to have AWS certifications (e.g., Solutions Architect, DevOps Engineer, SysOps Administrator). Experience designing or supporting disaster recovery and business continuity strategies. Familiarity with security compliance frameworks and implementing security best practices in cloud environments. Experience with serverless architectures using AWS Lambda, SAM, or the Serverless Framework. Experience supporting distributed engineering teams across multiple time zones. Exposure to data pipeline infrastructure or platforms used for large-scale data processing. FinOps certification or experience with cloud financial management and cost optimization practices. Base Salary Compensation Range $90,489.00-132,711.00 Incentive Target Percentage 12.5% Annual Morningstar's hybrid work environment gives you the opportunity to collaborate in-person each week as we've found that we're at our best when we're purposely together on a regular basis. In most of our locations, our hybrid work model is four days in-office each week. A range of other benefits are also available to enhance flexibility as needs change. No matter where you are, you'll have tools and resources to engage meaningfully with your global colleagues. 100_MstarResCanad Morningstar Research, Inc. (Canada) Legal Entity

What you’ll do

Design and maintain AWS cloud infrastructure and CI/CD pipelines to ensure platform stability and security. Lead reliability initiatives including disaster recovery, monitoring automation, and collaboration with engineering teams to embed SRE best practices.

Requirements

Requires 5+ years of experience in SRE or DevOps with strong proficiency in AWS, Infrastructure as Code, and containerization. A Bachelor's degree in Computer Science or a related field is expected along with scripting skills in Python or Bash.

Benefits

  • Incentive Target Percentage
  • Hybrid Work Environment

Listed skills

  • GitHub · Preferred
  • CI/CD · Preferred
  • Reliability · Preferred
  • Collaboration · Preferred
  • Docker · Preferred
  • Teams · Preferred
  • Compliance · Preferred
  • AI-assisted development · Preferred
  • Development · Preferred
  • Amazon Web Services · Preferred
  • Claude · Preferred
  • Python · Preferred

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • AWS
  • Terraform
  • CI/CD
  • Docker
  • Kubernetes
  • Python
  • Bash
  • Splunk
  • New Relic
  • SRE Principles
  • Linux Administration
  • Infrastructure as Code
  • CloudWatch
  • Incident Management
  • Capacity Planning
  • AI-assisted Development

Job areas

  • Technology
  • Software
  • Engineering
  • Data & Analytics
  • Finance & Accounting

More jobs from Morningstar

See all jobs from Morningstar