Transform operations at TJX Canada as a Site Reliability Engineer. Leverage AIOps and automation to enhance system reliability and incident response across platforms.
As an SRE with over six years' experience, you will drive initiatives that improve observability and automate manual tasks. Collaborate in a vibrant team workplace that encourages innovation while working with cutting-edge AI technologies and operational excellence across diverse platforms like Azure Cloud.
Key Responsibilities:
• Implement SRE best practices and process improvements
• Build automation scripts and self-healing workflows
• Support AI-driven capabilities like anomaly detection
• Enhance observability with logs, metrics, and dashboards
• Participate in incident response and root cause analysis
Requirements:
• 6+ years in SRE, DevOps, or related roles
• Proficiency with APM tools like Splunk and Datadog
• Solid scripting skills in Python, Java, or PowerShell
• Experience with Azure cloud settings
• Knowledge of distributed systems and APIs
Drive operational excellence using automation and AI technologies to make a positive impact on critical business applications.
J-18808-Ljbffr
📌 Site Reliability Engineer At Tjx Canada Ontario
🏢 The TJX Companies
📍 Canada
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.