Oracle: Site Reliability Engineer (Toronto)

Oracle: Site Reliability Engineer (Toronto)

12 Sep
|
Pager
|
Toronto

12 Sep

Pager

Toronto

NYSE: PD) is the global leader in AI-first digital operations. By automatically detecting, diagnosing, and remediating issues, the PagerDuty Platform orchestrates AI agents and automated workflows with context from over 750 integrations. Notable customers include Chipotle, Cloudflare, Docusign, Fox, Nvidia, Salesforce, Spotify, Zoom and more.

We are growing rapidly and hiring top talent with leading AI skills across engineering, sales, product, marketing, and beyond as we build the leading digital operations platform. As a Site Reliability Engineer II on the Core Infrastructure team in our Atlanta office, you'll help build and operate the foundational infrastructure that powers PagerDuty's real-time digital operations platform. daily, enabling customers to detect, respond to, and resolve incidents quickly and You'll work at the intersection of platform evolution and operational excellence, building and evolving foundational network, compute, and ingress infrastructure while scaling and security of the services our customers rely on to keep their businesses running as PagerDuty continues to grow across products, regions, and customer use cases. Support and improve foundational infrastructure, including networking, compute platforms, Kubernetes clusters, and ingress/traffic management systems.

Participate in agile rituals (standups, planning, retros) and communicate progress/risks early You stay current on technical trends to suggest innovative tools and approaches to interesting problems Monitor system health using metrics, logs, and alerts,



and participate in 24/7 on-call rotations to help detect, respond to, and resolve incidents. 3+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles ~ Hands-on experience operating Linux-based systems in production environments ~ Working knowledge of networking fundamentals, such as load balancing, DNS, TLS, and ingress traffic flow ~ AWS, GCP, Azure), including networking and compute concepts ~ Proficiency in at least one programming language (e.g., Python, Ruby, Go, etc.) ~ Experience with AWS cloud networking concepts such as VPCs, subnets, routing, security groups, and load balancers EKS), including cluster upgrades, networking, or ingress configuration Experience with monitoring, observability, and logging platforms (e.g., Our values guide how we support customers, collaborate with colleagues, develop products, and foster a culture of belonging.

People

Leaders at PagerDuty are responsible for creating high performance environments that drive accountability. Each dimension has three associated competencies to give leaders a shared language for guiding their development, career, promotion, and succession planning discussions.

Our Manager

Expectations serve as a practical guide for managers to understand their responsibilities, prioritize their efforts, and drive engagement and performance. As a global organization, our total rewards approach is competitive with industry standards and aligned with local laws and regulations. Learn more, including country-specific offerings, on our advantages site .

📌 Oracle: Site Reliability Engineer (Toronto)
🏢 Pager
📍 Toronto

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: oracle: site reliability engineer (toronto) / toronto

Subscribe to this job alert:

Get the latest job offers by email for: oracle: site reliability engineer (toronto) / toronto