01 Aug
|
Tru India
|
Toronto
Join our team as a Site Reliability Engineer focused on building and supporting our next-generation AI platform. This hybrid role, located in the Greater Toronto Area, allows you to enhance system reliability through advanced cloud strategies. As a key player in this position, you will deploy and manage infrastructure and containerized applications on Microsoft Azure.
Your role will involve improving operational visibility through dashboards, scaling applications intelligently, and implementing monitoring solutions to optimize performance. Your contributions will directly affect platform reliability in preparation for increased demand in 2027. Key Responsibilities:
- Deploy, configure, and maintain Azure cloud infrastructure
- Oversee container lifecycle and health monitoring
- Manage scaling policies for applications and resources
- Design and develop operational dashboards using relevant tools
- Collaborate on monitoring and alerting for infrastructure services Requirements:
- Minimum 5 years in a DevOps or SRE role
- Advanced knowledge of Azure and Kubernetes deployments
- Expertise in containerization with Docker
- Proficient in CI/CD pipelines and monitoring systems
- Solid experience in troubleshooting cloud-based applications Make an impact on our AI initiatives by enhancing system performance and reliability!
📌 Site Reliability Engineer - Hybrid Opportunity (Toronto)
🏢 Tru India
📍 Toronto