11 Sep
|
Hays
|
Vancouver
Job Title: SRE Architect
Location: Vancouver, BC (hybrid)
Duration: Long term contract
Rate: CAD75-CAD85/hr
/ Responsibility
- Set the project-wide reliability direction and embed SRE principles throughout solution design, delivery, release, and production operations.
- Establish service objectives, engineering standards, availability expectations, and go-live readiness criteria based on business criticality.
- Assess architecture and operational processes for failure scenarios, scalability constraints, capacity requirements, recovery capabilities, and production risks.
- Shape monitoring and service health visibility to enable teams to detect customer impact early and respond through actionable alerts and diagnostics.
- Lead major incident governance, root cause reviews, corrective action tracking, and the prevention of recurring issues.
- Identify repetitive operational tasks and drive engineering improvements that simplify delivery, recovery, and day-to-day support.
- Coordinate onsite and offshore teams by aligning plans, clarifying ownership, managing dependencies, and ensuring effective handoffs across time zones.
- Serve as the primary stakeholder-facing SRE lead for technical decisions, governance meetings, escalations, progress reporting, and continuous improvement initiatives.
- Coach engineers and influence cross-functional teams to make consistent, risk-aware decisions across the project.
Must have skills:
- Reliability architecture and production engineering.
- Distributed systems, cloud operations, and engineering automation.
- Technical leadership, consulting, and client communication.
Additional Information SRE mindset: Takes end-to-end ownership, uses evidence to prioritize work, anticipates operational risk, favors sustainable engineering over short-term fixes, and continuously improves customer outcomes.
Role scope: Operates across multiple engineering and support groups, provides functional direction rather than relying on formal reporting lines, and represents the reliability perspective in project and customer forums.
Key outputs: Reliability roadmap, service-level framework, architecture assessment findings, production-readiness decisions, health-view standards, risk register, improvement backlog, and governance reporting.
Role differentiator: Connects deep engineering judgement with business context, turning complex operational concerns into practical decisions, accountable actions, and measurable value.
Preferred background: Significant experience supporting business-critical, distributed production environments and advising senior technical stakeholders. Relevant industry certifications are advantageous.
Team culture: Transparent, cooperative, curious, and blameless; encourages constructive challenge, shared learning, clear accountability, and prevention-focused action after incidents.
📌 SRE Architect (Vancouver)
🏢 Hays
📍 Vancouver