Site Reliability Engineer (Vancouver)

Site Reliability Engineer (Vancouver)

06 Sep
|
PayByPhone
|
Vancouver

06 Sep

PayByPhone

Vancouver

About PayByPhone At PayByPhone, our strength is our people. Behind our product is a talented, creative, and driven multi-disciplinary team united by a shared ambition: to make everyday mobility simpler. We believe innovation should be collaborative, learning should be constant, and work should be enjoyable. As we grow, we’re looking for people who want to grow with us.

Together, we’re on an ambitious mission to create intuitive technology solutions that deliver world-class user experiences. We are a fast-growing, forward-thinking company and already help more than 60 million users across North America and Europe. Our technology helps millions of consumers pay quickly, easily, and securely — without waiting in line, carrying change, or worrying about costly fines.

About The Role Location: Vancouver, BC

Employment type: Full-Time, Permanent

Reports to: Site Reliability Lead, Platform Engineering

We're looking for Site Reliability Engineers to help build and operate highly available, secure, and scalable SaaS systems at PayByPhone. This role is open across the experience spectrum — whether you're a few years into an SRE or ops career or a seasoned specialist, we'll shape scope, mentorship, and support around where you're starting from.

This role calls for a motivated, quality- and results-oriented person who enjoys collaborating with cross-functional teams of skilled developers. The focus of the role is to:

Help ensure the PayByPhone platform meets its availability and stability requirements

Drive continuous improvements in our software quality assurance processes, practices, and culture

Working under the Site Reliability Lead (SRL), this role helps drive reliability by ensuring deliverables are implemented on time and on budget, and are operating at the expected levels of 24/7/365 at four nines of availability (target 99.99%) and within specific customer SLAs. This role also supports the SRL's agenda of incident prevention, improved incident response, and timely remediation

Key Responsibilities Help ensure platform reliability and stability deliverables are implemented on time and on budget, and are operating at 24/7/365 at four nines of availability, within specific customer SLAs

Standardize our quality assurance plans and templates for cross-team projects

Assist in quality assurance tooling selection and operational procedures

Help ensure continuity across all critical business transactions

Generate and report on reliability metrics to various stakeholders

Work collaboratively as a member of the team to define, refine, and execute the Platform Reliability / DevOps and SRE roadmap

Support PayByPhone in meeting security and compliance requirements by collaborating with the Security & Compliance team and implementing security and compliance tooling, processes, and policies as part of the CI/CD process

Work collaboratively with release & support management leaders on compliant and secure processes for triage, investigation, resolution,



and release of software

Assist the SRL in the quality assurance process within the SDLC, including automation language selection and usage

Help build a solid sense of ownership and accountability across teams and individual contributors, reflected at the code level and in implementation and operations

Contribute to high-quality execution and technical and operational excellence

Gather and analyze metrics from operating systems and applications to support performance tuning and fault-finding

Participate in system design consulting, assisting the Platform team where needs overlap with reliability and availability

Assist in scoping, developing, and testing the ongoing disaster recovery plan under the SRL's leadership

Help own observability, monitoring, and alerting tools under the SRL's leadership — proactively looking for ways to improve, standardize, document, and train

Provide operational support: documentation and debugging of production issues, including — Being available to join Sev-1 pages

Assisting the SRL with responsibilities such as post-mortems and on-call bootcamps Assist with CI/CD toolset setup and support (GitLab runners) from a reliability standpoint

Support cost reduction, budget setting, and monitoring across the platform and the tools used within it

Support software reliability practices (logging standards, use of reliable libraries, SLA/SLO goals)

Participate in on-call responsibilities when needed

Maintain a personal data plan to support your on-call responsibilities

Key Requirements 3+ years of experience in software development, delivery, or site reliability / operations for large, complex software systems spanning both legacy and modern stacks. Equivalent experience gained through non-traditional paths is welcome.

Bachelor's or higher degree in Computer Science, Computer Engineering, or a related technical field is preferred; equivalent hands-on experience will also be considered.

Experience working with high-performing SRE, Ops, or Dev teams

Solid grounding in quality assurance discipline, software quality management, and related frameworks and tools

Working experience with Amazon Web Services (AWS) solution architectures and technologies

Experience building verification and validation practices into end-to-end delivery pipelines, from business development through the delivery phase, to accelerate product launch to market

Familiarity with testing techniques such as unit, functional requirement, performance, GUI, regression, integration, system load, vulnerability assessment, security testing, and test automation





Understanding of SaaS multi-tenant and distributed / micro-service architectures

Understanding of DevOps principles, processes, and tools (e.g. IaC, CI/CD, and orchestration)

Understanding of cloud computing architecture, services, and platforms

Understanding of web and/or mobile development technologies and programming/scripting languages

Ability to program (structured and OOP) using one or more high-level languages, such as Python or JavaScript/TypeScript

Experience with distributed storage technologies such as NFS, HDFS, and Amazon S3, as well as dynamic resource management frameworks

Experience with Infrastructure as Code (IaC)

A proactive approach to identifying problems, performance bottlenecks, and areas for improvement

Comfortable working with and supporting cross-functional teams

Strong written communication, including technical documentation and training materials

What We Offer Compensation: The expected salary range for this role is $110,000 – $120,000 CAD. Final compensation will be based on factors such as experience, skills, qualifications, and internal equity.

Retirement Savings Program: Access to our retirement savings program (RRSP for Canada / 401(k) for U.S.-based employees).

Vacation: All permanent full-time employees start with 4 weeks of vacation per year.

Work from Anywhere: Up to 15 days of work from anywhere subject to management and IT Security approval.

Personal Days: We provide 5 personal days annually, in addition to paid sick days, to support flexibility and work-life balance.

Comprehensive medical & dental coverage:

Employee Assistance Program (EAP): Access to confidential support services and resources for you and your family.

Career Growth & Learning Support: Opportunities for professional development, continuous learning, and career progression.

Working at PayByPhone We operate in a world that’s constantly evolving — and change is something we embrace. Our values guide how we show up for one another and for our customers every day. In short, we:

Make things happen

Stay curious

Work together

Have fun

See through our customers’ eyes

These principles shape how we collaborate, innovate, and deliver on our commitments.

We’re also committed to fostering a diverse and representative workforce and an inclusive environment where everyone is treated with respect and fairness. We do not tolerate discrimination or harassment in our workplace or throughout our hiring process. Our hiring decisions are grounded in business needs, role requirements, and individual qualifications — ensuring we reflect the talent and communities we serve.

PayByPhone is committed to providing accommodation throughout the recruitment process. If you require accommodation, please reach out to us at [email protected].

Want to see our values in action? Visit our Instagram and LinkedIn. Curious about the story behind our values? Head over to our About Us page to learn more.

📌 Site Reliability Engineer (Vancouver)
🏢 PayByPhone
📍 Vancouver

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: site reliability engineer (vancouver) / vancouver