Elevate your career as a Senior Site Reliability Engineer with a focus on designing and maintaining scalable systems while influencing architectural decisions. This remote role emphasizes automation and on-call incident management.
In this position, you will utilize your 5+ years of experience in Site Reliability Engineering or Dev Ops to improve system performance and reliability. The role requires solid programming skills in languages like Rust, Go, or Python, alongside deep knowledge of AWS services and Kubernetes.
Collaborate with development teams and automate key processes to enhance observability and operational efficiency.
Key Responsibilities:
Design and maintain fault-tolerant distributed systems
Automate operational tasks and CI/CD pipelines
Lead blameless root cause analyses
Monitor and optimize system performance and latency
Build internal tools for self-service observability
Requirements:
Bachelor's degree in Computer Science or related field
5+ years in Site Reliability Engineering or Dev Ops
Proficient in AWS services and Kubernetes
Robust programming skills in high-level languages
Familiar with Linux/Unix systems and distributed architecture
Shape the future of reliable systems with your programming and AWS expertise in this Senior Site Reliability Engineer role.
J-18808-Ljbffr
📌 Senior Site Reliability Engineer Remote Toronto
🏢 Jobtailor
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.