12 Aug
|
Gemini Solutions]
|
Toronto
12 Aug
Gemini Solutions]
Toronto
Elevate platform reliability with us as a Senior Site Reliability Engineer. Own mission-critical systems and boost performance, scalability, and observability to achieve operational excellence.
We seek an experienced Senior Site Reliability Engineer to lead platform reliability initiatives. In this role, you will collaborate with engineering and business stakeholders to ensure systems are productive and highly available. Focus on automating operations and implementing best practices in reliability and performance.
Key Responsibilities:
Own performance and availability of production systems
Define and implement SLIs, SLOs, and error budgets
Lead incident response and service restoration
Enhance monitoring, logging, and alerting frameworks
Automate operational workflows to optimize efficiency
Requirements:
3+ years in SRE or DevOps roles
Experience in production-critical settings
Proficiency in Python for automation tasks
Skilled in AWS, Azure, or GCP
Familiarity with monitoring tools like Datadog or Prometheus
Take the lead in ensuring system reliability while driving innovation through automation.
J-18808-Ljbffr
📌 Senior Site Reliability Engineer Role Toronto
🏢 Gemini Solutions]
📍 Toronto