IBM is seeking an experienced Site Reliability Engineer to drive proactive reliability improvements across a multi-cloud platform (AWS, GCP, Azure). You’ll own tooling, post-mortem coaching, and incident response enhancements in a follow-the-sun team within Cloud Architecture and Reliability — Supportability.
Responsibilities include designing SLO/SLA frameworks, partnering with engineering leaders, and advancing observability and release processes.