Senior Site Reliability/Observability Engineer (Toronto)

Senior Site Reliability/Observability Engineer (Toronto)

05 Aug
|
Kyndryl
|
Toronto

05 Aug

Kyndryl

Toronto

Job Title: Senior Site Reliability/Observability Engineer

Duration: 3-Month Contract (Aug - October)

Location: Toronto, ON (Hybrid - 2 to 3 days onsite per week)

Job Overview

We are seeking a Senior Site Reliability/Observability Engineer to support a critical operational excellence initiative focused on Application Performance Monitoring (APM), Operational Readiness, Compliance Automation, and Dashboarding.

The successful candidate will be responsible for designing, developing, and implementing monitoring solutions, audit workflows, compliance automation, and executive reporting dashboards that improve system visibility, operational governance, and business decision-making.

This is a highly hands-on role requiring strong full-stack development experience, integration expertise, and the ability to work closely with operations, compliance, and business stakeholders.

Responsibilities:

Application Performance Monitoring (APM)

- Design and implement APM solutions to monitor application health and performance
- Configure monitoring for latency, throughput, error rates, and user experience metrics
- Establish performance baselines and alerting thresholds
- Enable end-to-end transaction tracing and dependency mapping

Operational Readiness Framework

- Develop pre-deployment validation processes and operational readiness checklists
- Create operational runbooks, standards, and support documentation
- Implement deployment validation and approval workflows

Audit & Compliance Automation

- Design and implement automated audit trails for system changes and deployments




- Build compliance workflows for evidence collection and reporting
- Develop audit dashboards and compliance reporting capabilities

Dashboarding & Visualization

- Build role-based dashboards for executives, operations, and technical teams
- Provide real-time visibility into system health, incidents, and performance trends
- Enable self-service reporting and operational insights

Qualifications

Must-Have Skills

- 8+ years of Full Stack Development experience
- Strong experience building enterprise dashboards and reporting solutions
- Experience implementing APM/Observability platforms (Dynatrace, AppDynamics, Datadog, New Relic, Splunk, etc.)
- Robust experience with API development and system integrations
- Experience building workflow automation and operational tooling
- Front-end development experience (React, Angular, or similar)
- Back-end development experience (.NET, Java, Node.js, Python, or similar)
- Experience with SQL and data visualization technologies
- Strong understanding of DevOps, monitoring, and operational best practices

Nice-to-Have Skills

- Experience with compliance, audit, or governance automation
- Experience building executive and operational dashboards
- Knowledge of ITIL, operational readiness, or change management processes
- Cloud experience (Azure, AWS, or GCP)
- Experience with CI/CD pipelines and Infrastructure as Code

#IndKyn

- Please note this is for a contract position with one of our clients and not a fulltime employment role with Kyndryl Canada**

#J-18808-Ljbffr

📌 Senior Site Reliability/Observability Engineer (Toronto)
🏢 Kyndryl
📍 Toronto

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability/observability engineer (toronto) / toronto

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability/observability engineer (toronto) / toronto