05 Oct
|
XYZ Venture Capital
|
Toronto
05 Oct
XYZ Venture Capital
Toronto
At Rootly, we are on a mission to be the go‑to way companies respond when things go wrong, helping every organization be more reliable. We do this by building an industry‑leading incident management platform that allows companies around the world consistently and quickly resolve incidents. Customers love Rootly.
See why our customers have reviewed us 5 stars on G2. We conduct monthly financial reviews as a team so everyone has a pulse on the health of the business and publish what we are building in our weekly changelog. This is an opportunity to join Rootly as an early SRE leader and shape our technical foundation.
What you’ll be doing one day could look very different the next. You will be empowered to identify opportunities that will help us grow and own it. In short, this role is designed for individuals that crave ownership, stimulating technical challenges, love shipping fast, and are mission‑driven.
Embed with product teams to enhance observability, reliability, and performance of their services. Own our CI/CD pipelines, observability tooling, monitoring systems, and incident response processes. Build tools and automation to eliminate manual toil, improve engineering velocity and developer experience, and improve system reliability.
Collaborate deeply across engineering to understand systems at the code level and surface cross‑cutting reliability, performance, and scaling concerns. Architect and scale our infrastructure, ensuring best‑in‑class performance, availability, and operational excellence. Drive capacity planning efforts to ensure our infrastructure is resilient and scalable as we grow.
Define and manage SLOs and error budgets in partnership with Engineering teams who own production services. Be vocal – act as a strong voice and force of reliability, quality, performance, and scalability. 5+ years of experience in an SRE, Platform, or Infrastructure Engineering role. ~5+ years of experience writing software in a production environment. ~ Robust technical knowledge of cloud infrastructure, distributed systems, and reliability practices. ~ Strong understanding of observability, performance tuning, and scaling strategies. ~ Deep familiarity with incident response, monitoring, and CI/CD systems. ~ Hands‑on experience supporting web or RPC services at meaningful scale. ~ not shell scripts alone, but production‑grade software.
Experience with Ruby and Go is a plus. We’re building something category‑defining and want teammates who crave ownership, love solving hard problems, and thrive in a high‑bar, high‑impact environment. Competitive compensation and early equity in a fast‑growing, venture‑backed company. ~ Comprehensive medical, dental, and vision coverage. ~3 weeks of vacation, plus unlimited sick and mental health days, and a company‑wide end‑of‑year shutdown to recharge. ~$500 stipend for home office setup. ~ Unlimited token usage and access to AI tools. ~ We aim to create an environment where every team member at Rootly feels like they belong so they can have a greater impact on our business and customers.
We do not discriminate on the basis of race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. #
📌 Senior Site Reliability Engineer (Remote) (Toronto)
🏢 XYZ Venture Capital
📍 Toronto