Manager, Site Reliability Engineering (Toronto)

Manager, Site Reliability Engineering (Toronto)

28 Aug
|
Akkodis
|
Toronto

28 Aug

Akkodis

Toronto

Title:

Manager, Site Reliability Engineering (SRE)

Position Type:

Full-time, Permanent

Location:

Etobicoke, ON

Working Model:

Hybrid - 3 days per week in office

Compensation:

$140K - 165K per annum plus bonus and perks

About Our Client Our client is a leading Canadian organization operating highly available, mission-critical technology platforms that support millions of transactions and customer interactions annually.

As part of a broader technology transformation and operational excellence initiative, the organization is continuing to invest in its Site Reliability Engineering (SRE) capabilities to improve reliability, scalability, observability, and production resilience across a large portfolio of business-critical applications.

This is an exciting opportunity to join a growing SRE function at an important stage of its evolution, helping shape reliability practices, operational standards, automation strategies, and engineering culture.

About the Opportunity We are seeking a

hands-on

Manager, Site Reliability Engineering (SRE)

to lead a team of SRE engineers responsible for improving the reliability, availability, performance, and operational excellence of critical enterprise applications and platforms.

This role offers a balance of

people leadership and technical leadership , making it ideal for an experienced SRE, Platform Engineering, DevOps, or Production Engineering leader who enjoys building high-performing teams while remaining engaged in technical strategy and operational improvements.

You will partner closely with Application Development, DevOps, Infrastructure, Security, and Incident Management teams to drive observability, automation, incident response, resiliency, and continuous reliability improvements across both cloud and on-premises environments.

What You'll Do

Lead and develop a team of Site Reliability Engineers

Drive adoption of SRE principles, including SLOs, SLIs, error budgets, and reliability best practices

Improve observability, monitoring, alerting, and operational health across critical services





Reduce operational toil through automation and self-healing capabilities

Oversee production support, incident management, escalation processes, and on-call operations

Partner with engineering teams to embed reliability into application design and delivery pipelines

Lead capacity planning, resiliency testing, and disaster recovery readiness initiatives

Act as a senior escalation point during major incidents

Improve deployment reliability and reduce operational risk

Establish runbook standards, documentation practices, and operational governance

Collaborate with leadership teams to execute strategic SRE initiatives and roadmaps

Foster a blameless culture focused on continuous improvement and operational excellence

What You Bring Must-Have Qualifications

8+ years of experience supporting large-scale production environments, distributed systems, or cloud platforms

3+ years of people leadership experience managing technical teams

Experience leading SRE, Platform Engineering, DevOps, Production Operations, or Reliability Engineering teams

Strong understanding of Site Reliability Engineering principles and practices

Experience with observability and monitoring platforms such as Dynatrace, Datadog, New Relic, AppDynamics, or similar

Hands‑on experience with Azure (preferred) or AWS

Strong Kubernetes and container platform experience

Expertise with Infrastructure as Code and automation tools such as Terraform and Ansible

Experience leading incident response, problem management, and production operations

Strong scripting or programming skills (Python, PowerShell, Bash, etc.)

Excellent communication, stakeholder management,



and leadership skills

Nice-to-Have Qualifications

Experience in payments, fintech, banking, or highly regulated environments

Knowledge of PCI-DSS, NIST, or related compliance frameworks

Experience with change management and release governance processes

Experience implementing SLOs, SLIs, and error budget frameworks

Background building or maturing an SRE function

Experience supporting enterprise-scale cloud transformation initiatives

What You'll Love About This Opportunity

Opportunity to help shape and mature a growing SRE organization

Lead a team responsible for business-critical services and platforms

Drive meaningful improvements in reliability, observability, and automation

Balance people leadership with hands‑on technical influence

Work alongside highly skilled Engineering, Infrastructure, Security, and DevOps teams

High visibility role with broad organizational impact

Collaborative culture focused on innovation, resilience, and continuous improvement

Competitive compensation, bonus, and benefits package

Strong growth opportunities opportunities within a modern technology organization

Significant We thank all applicants for their interest in this opportunity. Only candidates meeting the above qualifications will be contacted for further discussions.

Akkodis Canada will never share your resume or any personal details without your explicit consent.

We thank all applicants for their interest in this opportunity. Only candidates meeting the above qualifications will be contacted for further discussions.

Our Commitment: At Akkodis, part of The Adecco Group, our purpose is simple: to make the future work for everyone. We live our values, Passion, Collaboration, Inclusion, Courage, and Customers at Heart, by fostering a workplace where diversity is celebrated and every voice matters. We encourage applications from individuals of all backgrounds and identities. Together, we’re making the future work for everyone.

#J-18808-Ljbffr

📌 Manager, Site Reliability Engineering (Toronto)
🏢 Akkodis
📍 Toronto

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: manager, site reliability engineering (toronto) / toronto

Subscribe to this job alert:

Get the latest job offers by email for: manager, site reliability engineering (toronto) / toronto