Drive operational excellence as a Senior Site Reliability Engineer with a focus on AIOps and automation. This role is critical for system reliability and efficiency.
With 3+ years in SRE or Systems Engineering, you will enhance system resilience and ensure proactive monitoring through cutting-edge AIOps solutions. Collaborate with cross-functional teams while taking ownership of projects and utilizing your deep knowledge of observability tools and automation technologies.
Key Responsibilities:
• Develop and maintain AIOps and SRE capabilities
• Oversee disaster recovery and resiliency initiatives
• Implement observability and logging frameworks
• Enhance infrastructure platform support
• Collaborate on compliance and governance activities
Requirements:
• 3+ years of relevant engineering experience
• Experience with ServiceNow and ITSM techniques
• Skilled in monitoring and observability tools
• Background in containerized platforms like Kubernetes
• Familiar with scripting in Python or PowerShell
Leverage your skills to ensure system reliability and contribute significantly to our operational strategies.
#J-18808-Ljbffr