Site Reliability Engineer
Design and optimize monitoring and observability solutions using Dynatrace to ensure application reliability. Support AWS-hosted environments by resolving performance bottlenecks and automating operational tasks.
- On-site
- Toronto, ON
- Posted Jul 28, 2026
- Apply by Aug 27, 2026
- 1 position
Job summary
We are looking for an experienced Site Reliability Engineer (SRE) to join a leading financial services client. This role is ideal for someone with strong Dynatrace expertise who enjoys improving application reliability, monitoring, automation, and operational excellence in cloud environments. Key Responsibilities Design, implement, and optimize monitoring and observability solutions using Dynatrace. Monitor application and infrastructure performance, proactively identifying and resolving performance bottlenecks. Support production environments by investigating incidents, performing root cause analysis, and implementing permanent solutions. Develop dashboards, alerts, and monitoring strategies to improve system reliability and availability. Collaborate with development, DevOps, and infrastructure teams to enhance application resilience. Automate operational tasks and continuously improve system performance and operational efficiency. Support AWS-hosted applications and cloud infrastructure. Required Skills Strong hands-on experience with Dynatrace (monitoring, dashboards, alerting, troubleshooting, and performance analysis). Experience working in Site Reliability Engineering (SRE) or Production Support environments. Knowledge of AWS cloud services and application deployment. Experience with incident management, problem management, and root cause analysis. Strong scripting or automation skills (Shell, Python, or similar) are an asset. Excellent troubleshooting and communication skills. Nice to Have Previous software development experience (Java, Python, or other programming languages). Experience supporting cloud-native or microservices-based applications. Familiarity with CI/CD pipelines and DevOps practices. Experience in the banking or financial services industry.
What you’ll do
Design and optimize monitoring and observability solutions using Dynatrace to ensure application reliability. Support AWS-hosted environments by resolving performance bottlenecks and automating operational tasks.
Requirements
Requires strong hands-on experience with Dynatrace, AWS cloud services, and SRE or production support environments. Proficiency in scripting languages like Python or Shell and experience with incident management are essential.
Other relevant skills
Identified from the job description. Confirm important requirements above.
- Dynatrace
- Site Reliability Engineering
- AWS
- Incident Management
- Root Cause Analysis
- Shell Scripting
- Python
- Monitoring
- Observability
- Automation
- Production Support
- Cloud Infrastructure
- CI/CD
- DevOps
- Microservices
- Java
Job areas
- Technology
- Software
- Engineering
- Finance & Accounting
Additional details
- Minimum experience
- 5+ years
- Apply by
- Aug 27, 2026
- Posting language
- English
- Working hours
- 40 hours per week
- Seniority
- Mid-Senior level
- Application method
- Direct apply is available
