Join SAP as a Site Reliability Engineer in Montreal, where you'll support business-critical cloud services. Contribute to monitoring, troubleshooting, and enhancing infrastructure reliability in this hybrid position. In this role, you will integrate into the established Site Reliability Engineering team for the SAP Business AI Platform.
Your responsibilities will include investigating service incidents, driving root cause analyses, and developing software solutions. Proficiency in Kubernetes and cloud platforms like AWS or Azure is essential, along with excellent communication skills. Key Responsibilities:
- Operate and support critical cloud services
- Perform deep technical investigations during live incidents
- Drive improvements to prevent service outages
- Build tools for monitoring and troubleshooting
- Participate in on-call rotations for major incidents
Requirements:
- 2+ years experience in SRE
- Proficiency in Kubernetes and container technologies
- Experience with Unix/Linux systems
- Solid scripting and CI/CD skills
- Fluency in English
Become a key player in ensuring the reliability of SAP's cloud services through exceptional monitoring and analysis.
📌 Site Reliability Engineer at SAP (Montreal)
🏢 SAP
📍 Montreal