Elevate the reliability of voice AI infrastructure with our Retail and Restaurants AI team as a Senior Site Reliability Engineer. This role emphasizes Google Cloud Platform and automation.
As part of your position, you will ensure optimal performance and scalability of systems supporting millions of interactions. You will leverage 12+ years of experience in Site Reliability Engineering to architect automation strategies and robust monitoring solutions. Your ability to partner across teams will enhance operational effectiveness.
Key Responsibilities:
• Maintain highly available infrastructure on Google Cloud Platform
• Drive automation of CI/CD pipelines for deployment efficiency
• Implement observability strategies to identify system challenges
• Collaborate on performance and reliability improvements for services
• Oversee compliance initiatives for security standards
Requirements:
• 12+ years in software engineering, focusing on SRE or DevOps
• Expert in Google Cloud Platform services, including GKE
• Skilled with Infrastructure as Code tools like Pulumi
• Extensive experience with Kubernetes and monitoring tools
• Robust problem-solving and communication capabilities
Utilize your leadership skills to streamline AI infrastructure with our dedicated team.
#J-18808-Ljbffr