09 Aug
|
Gemini Solutions
|
Ontario
09 Aug
Gemini Solutions
Ontario
Elevate platform reliability with us as a Senior Site Reliability Engineer. Own mission-critical systems and boost performance, scalability, and observability to achieve operational excellence.
We seek an experienced Senior Site Reliability Engineer to lead platform reliability initiatives. In this role, you will collaborate with engineering and business stakeholders to ensure systems are effective and highly available. Focus on automating operations and implementing best practices in reliability and performance.
Key Responsibilities:
• Own performance and availability of production systems
• Define and implement SLIs, SLOs, and error budgets
• Lead incident response and service restoration
• Enhance monitoring, logging, and alerting frameworks
• Automate operational workflows to optimize efficiency
Requirements:
• 3+ years in SRE or DevOps roles
• Experience in production-critical environments
• Proficiency in Python for automation tasks
• Skilled in AWS, Azure, or GCP
• Familiarity with monitoring tools like Datadog or Prometheus
Take the lead in ensuring system reliability while driving innovation through automation.
#J-18808-Ljbffr
📌 Senior Site Reliability Engineer Role (Ontario)
🏢 Gemini Solutions
📍 Ontario