05 Oct
|
Google Canada
|
London
05 Oct
Google Canada
London
Google is looking for a Staff Site Reliability Developer to join the Google Unified Security and Threat Operations team in Waterloo, Ontario. This is a senior-level engineering role sitting at the intersection of software development and systems engineering — where the problems are genuinely complex and the scale is unlike almost anywhere else in the industry.
As part of Google Cloud's Site Reliability Engineering (SRE) organization, you'll be responsible for the full lifecycle of large-scale, massively distributed systems — from initial design through deployment, live operations, and continuous improvement. The work spans everything from capacity planning and launch reviews to automation, incident response, and performance monitoring.
About the Role: Staff Site Reliability Developer
This role is squarely focused on keeping Google Cloud's services — both internal and customer-facing — reliable, performant, and continuously improving. You'll engage directly with system design, build software platforms and frameworks, and help scale infrastructure sustainably through automation. Much of the day-to-day work involves optimizing existing systems, building shared infrastructure, and eliminating toil so that engineering teams can move faster with confidence.
Google's SRE culture places a strong emphasis on intellectual curiosity, psychological safety, and blameless postmortems. You'll work alongside people with diverse backgrounds and perspectives, encouraged to think big, take calculated risks, and self-direct meaningful projects. The team also values mentorship and growth — both giving and receiving it.
Benefits and Salary
This position offers a competitive annual salary ranging from $216,000 to $221,000 CAD, along with a 20% bonus target, equity, and a comprehensive advantages package. For full details on Google's benefits offerings, visit Google's careers benefits page.
Job Details
Company: Google
Location: Waterloo, ON, Canada
Pay: $216,000 – $221,000 CAD/year + 20% bonus target + equity + benefits
Responsibilities
In this role, you'll be deeply involved in the entire lifecycle of Google Cloud services — not just keeping things running, but actively shaping how they're built and evolved. Your contributions will directly affect the reliability and velocity of systems used by Google's internal teams and external customers worldwide.
- Engage across the full service lifecycle — from inception and design through deployment, live operation, and ongoing refinement
- Support pre-launch activities including system design consulting, developing software platforms and frameworks, capacity planning, and launch reviews
- Maintain live services by measuring and monitoring availability, latency, and overall system health
- Scale systems sustainably through automation and evolve them by pushing for changes that improve reliability and development velocity
- Practise sustainable incident response and lead blameless postmortems to continuously improve operational processes
- Optimize existing systems and build infrastructure that eliminates repetitive work through smart automation
Requirements / Skills
Google is looking for an experienced engineer who brings deep expertise in distributed systems and software development, along with the judgment and communication skills to operate effectively at a senior level. Strong problem-solving instincts and a collaborative mindset are just as important as technical credentials.
- Bachelor's degree in Computer Science, a related field, or equivalent practical experience
- 8 years of software development experience in one or more programming languages
- 3 years of experience designing, analyzing, and troubleshooting distributed systems
- Master's degree in Computer Science or Engineering is preferred
- Strong proficiency in coding, algorithms, and complexity analysis at scale
- Demonstrated ability to work in a blameless, collaborative engineering culture
Bachelor's degree in Computer Science, a related field, or equivalent practical experience. Master's degree in Computer Science or Engineering preferred.
#J-18808-Ljbffr
📌 Staff Site Reliability Developer, Google Unified Security and Threat Operations (London)
🏢 Google Canada
📍 London