Elevate your career as a Senior Site Reliability Engineer with a focus on designing and maintaining scalable systems while influencing architectural decisions. This remote role emphasizes automation and on-call incident management.
In this position, you will utilize your 5+ years of experience in Site Reliability Engineering or Dev Ops to improve system performance and reliability. The role requires strong programming skills in languages like Rust, Go, or Python, alongside deep knowledge of AWS services and Kubernetes.
Collaborate with development teams and automate key processes to enhance observability and operational efficiency.
Key Responsibilities:
- Design and maintain fault-tolerant distributed systems
- Automate operational tasks and CI/CD pipelines
- Lead blameless root cause analyses
- Monitor and optimize system performance and latency
- Build internal tools for self-service observability
Requirements:
- Bachelor's degree in Computer Science or related field
- 5+ years in Site Reliability Engineering or Dev Ops
- Proficient in AWS services and Kubernetes
- Robust programming skills in high-level languages
- Familiar with Linux/Unix systems and distributed architecture
Shape the future of reliable systems with your programming and AWS expertise in this Senior Site Reliability Engineer role.
#J-18808-Ljbffr
📌 Senior Site Reliability Engineer Remote (Toronto)
🏢 Jobtailor
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.