

[8SN] Site Reliability Engineer (SRE) – UI/UX
Top Benefits
About the role
Company Description
We are Software Mind, an awesome team of engineers who are ready to ramp up any top-notch company’s projects! Our aim? To always be one step ahead. Become part of a multicultural company in constant growth with an excellent work environment certified by Great Place To Work!
About the Client
Our client is a leading enterprise software company building highly scalable cloud-native platforms used by organizations around the world. Their engineering teams focus on delivering reliable, secure, and high-performing services while embracing modern DevOps, Kubernetes, and cloud technologies. You will join a team responsible for ensuring the stability, reliability, and operational excellence of a critical UI service running in production. #LI-DNI Job Description About the Role We are looking for a Site Reliability Engineer (SRE) – UI/UX to support the deployment, operations, and ongoing maintenance of a production UI service running on Kubernetes. This role focuses on monitoring service health, troubleshooting production issues, investigating incidents, and ensuring reliable service delivery. You will work closely with engineering and client teams to support production operations and complete work based on a client-directed backlog. While this role supports a UI-based service, it is not a frontend development position. Working knowledge of Web Components is required to perform first-level debugging of UI-related issues, but deep frontend development expertise is not expected.
What You’ll Do
Support the deployment, operations, and ongoing maintenance of production services running on Kubernetes. Monitor service health, availability, and performance. Investigate and troubleshoot production incidents using logs, monitoring, and debugging tools. Perform log analysis and incident debugging using Splunk. Identify service issues and collaborate with engineering teams to support timely resolution. Participate in incident response and production support activities. Perform first-level debugging of UI-related issues involving Web Components. Support service reliability and continuous improvement initiatives. Assist with CI/CD pipelines and cloud-native application operations when needed. Work effectively within a client-directed backlog and established priorities.
Qualifications
Required Qualifications 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Support, or a related role. Hands-on experience supporting the deployment, operations, and ongoing maintenance of production services running on Kubernetes. Experience monitoring service health, troubleshooting production issues, and supporting service reliability. Proficiency with Splunk for log analysis and incident debugging. Experience participating in production incident response and root-cause analysis. Working knowledge of Web Components and the ability to perform first-level debugging of UI-related issues. Strong troubleshooting, analytical, and problem-solving skills. Experience collaborating with software engineering and cross-functional teams. Ability to work independently and effectively within a client-directed backlog. Excellent written and spoken English, at least B2 level.
Additional Information
Preferred Qualifications Experience supporting CI/CD pipelines. Familiarity with multi-tenant services. Experience with cloud-native application operations. Experience supporting high-availability enterprise or SaaS platforms. Familiarity with additional monitoring and observability tools. Experience with cloud platforms such as AWS, Azure, or GCP. Familiarity with container and deployment technologies such as Docker and Helm.
What We Offer
Competitive salary and laptop Professional development and training opportunities Work with cutting-edge cloud and container technologies Flexible work arrangements and collaborative team environment Impact on organization-wide digital transformation initiatives
Not the right fit? Search for [8SN] Site Reliability Engineer jobs in Montreal, Quebec, Canada
About Software Mind
Software Mind is a global digital transformation partner with operations throughout Europe, the US and LATAM. Driven by tech and empowered by people, we provide companies with software engineers and autonomous, cross-functional development teams who manage software life cycles from ideation to release and beyond.
For over 20 years we’ve been enriching organizations with the talent they need to boost scalability, drive dynamic growth and bring disruptive ideas to life. Our top-notch engineering teams combine ownership with leading technologies, including cloud, AI, data science and embedded software to accelerate digital transformations and boost software delivery.
A culture, driven by trust, that embraces openness, craves more and acts with respect enables our experts to create evolutive solutions that support scale-ups, unicorns and enterprise-level companies around the world.
Similar Jobs


[8SN] Site Reliability Engineer (SRE) – UI/UX
Top Benefits
About the role
Company Description
We are Software Mind, an awesome team of engineers who are ready to ramp up any top-notch company’s projects! Our aim? To always be one step ahead. Become part of a multicultural company in constant growth with an excellent work environment certified by Great Place To Work!
About the Client
Our client is a leading enterprise software company building highly scalable cloud-native platforms used by organizations around the world. Their engineering teams focus on delivering reliable, secure, and high-performing services while embracing modern DevOps, Kubernetes, and cloud technologies. You will join a team responsible for ensuring the stability, reliability, and operational excellence of a critical UI service running in production. #LI-DNI Job Description About the Role We are looking for a Site Reliability Engineer (SRE) – UI/UX to support the deployment, operations, and ongoing maintenance of a production UI service running on Kubernetes. This role focuses on monitoring service health, troubleshooting production issues, investigating incidents, and ensuring reliable service delivery. You will work closely with engineering and client teams to support production operations and complete work based on a client-directed backlog. While this role supports a UI-based service, it is not a frontend development position. Working knowledge of Web Components is required to perform first-level debugging of UI-related issues, but deep frontend development expertise is not expected.
What You’ll Do
Support the deployment, operations, and ongoing maintenance of production services running on Kubernetes. Monitor service health, availability, and performance. Investigate and troubleshoot production incidents using logs, monitoring, and debugging tools. Perform log analysis and incident debugging using Splunk. Identify service issues and collaborate with engineering teams to support timely resolution. Participate in incident response and production support activities. Perform first-level debugging of UI-related issues involving Web Components. Support service reliability and continuous improvement initiatives. Assist with CI/CD pipelines and cloud-native application operations when needed. Work effectively within a client-directed backlog and established priorities.
Qualifications
Required Qualifications 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Support, or a related role. Hands-on experience supporting the deployment, operations, and ongoing maintenance of production services running on Kubernetes. Experience monitoring service health, troubleshooting production issues, and supporting service reliability. Proficiency with Splunk for log analysis and incident debugging. Experience participating in production incident response and root-cause analysis. Working knowledge of Web Components and the ability to perform first-level debugging of UI-related issues. Strong troubleshooting, analytical, and problem-solving skills. Experience collaborating with software engineering and cross-functional teams. Ability to work independently and effectively within a client-directed backlog. Excellent written and spoken English, at least B2 level.
Additional Information
Preferred Qualifications Experience supporting CI/CD pipelines. Familiarity with multi-tenant services. Experience with cloud-native application operations. Experience supporting high-availability enterprise or SaaS platforms. Familiarity with additional monitoring and observability tools. Experience with cloud platforms such as AWS, Azure, or GCP. Familiarity with container and deployment technologies such as Docker and Helm.
What We Offer
Competitive salary and laptop Professional development and training opportunities Work with cutting-edge cloud and container technologies Flexible work arrangements and collaborative team environment Impact on organization-wide digital transformation initiatives
Not the right fit? Search for [8SN] Site Reliability Engineer jobs in Montreal, Quebec, Canada
About Software Mind
Software Mind is a global digital transformation partner with operations throughout Europe, the US and LATAM. Driven by tech and empowered by people, we provide companies with software engineers and autonomous, cross-functional development teams who manage software life cycles from ideation to release and beyond.
For over 20 years we’ve been enriching organizations with the talent they need to boost scalability, drive dynamic growth and bring disruptive ideas to life. Our top-notch engineering teams combine ownership with leading technologies, including cloud, AI, data science and embedded software to accelerate digital transformations and boost software delivery.
A culture, driven by trust, that embraces openness, craves more and acts with respect enables our experts to create evolutive solutions that support scale-ups, unicorns and enterprise-level companies around the world.