Join our team as a Lead Site Reliability Engineer focused on enhancing system reliability and scalability. Drive incident management and automate workflows to improve operational efficiency.This role targets a technically skilled engineer committed to maintaining high availability for production systems. You will collaborate closely with cross-functional teams to implement reliability practices and oversee incident management efforts. Build resilient systems while leveraging your cloud infrastructure expertise.Key Responsibilities:Drive continuous improvements in system resilienceLead root cause analysis and implement long-term fixesBuild and maintain CI/CD pipelinesTroubleshoot distributed systems effectivelyCommunicate incidents clearly to stakeholdersRequirements:3+ years of SRE, DevOps experienceStrong SQL skills for validation and troubleshootingHands-on experience with containerization technologiesFamiliarity with Infrastructure as Code approachesKnowledge of DNS and networking fundamentalsMake your mark on our mission-critical systems while embracing cutting-edge reliability solutions.#J-18808-Ljbffr
📌 Lead Site Reliability Engineer Opening (Toronto)
🏢 Gemini Solutions
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.