Senior Site Reliability Engineer, Node Platform (British Columbia)

Senior Site Reliability Engineer, Node Platform (British Columbia)

06 Aug
|
Framework Ventures
|
British Columbia

06 Aug

Framework Ventures

British Columbia

Your Impact
You will design and build the infrastructure primitives that define how Chainlink Decentralized Oracle Networks (DONs) scale across internal systems and the decentralized ecosystem.

You will help create the CRE (Kubernetes-based) control plane that enables:

Deterministic horizontal scaling of DONs

Safe and repeatable infrastructure expansion

Improved operational efficiency and scalability

You will develop the core infrastructure components, including Kubernetes Operators and scaling automation, that Product teams will adopt and then might later be distributed to external node operators to improve decentralized scaling.

This is not an operational support role. You will be building the systems that define how Chainlink scales while shaping the reliability, scalability, and decentralization of protocol-level services.

Requirements

6–9+ years in SRE / Platform / Infrastructure Engineering

Proven experience scaling Kubernetes in high-throughput production environments

Deep knowledge of:

Scheduler behavior

StatefulSets & persistent workloads

Autoscaling strategies (HPA, VPA, KEDA, custom scaling)

Resource management & performance tuning

Multi-cluster and multi-region architectures

Experience in diagnosing production failures at the cluster scale

Solid Terraform or Crossplane experience

GitOps workflows (ArgoCD / Flux) experience

CI/CD reliability experience

Automation-first mindset

AWS production experience

Proficiency in Go (strongly preferred) or equivalent systems language

Desired Qualifications

Experience with web3 concepts (e.g., blockchain node lifecycle,



forks, reorgs, or RPC issues)

Experience with oracle systems, token architectures, or decentralized services

Experience scaling stateful high-availability distributed systems

Experience building internal platform primitives

Experience implementing custom autoscaling logic

Experience designing SLO strategies and error-budget usage

Experience improving diagnosability and observability frameworks

Experience working in high-ambiguity environments

Experience operating blockchain infrastructure in production

Certified Kubernetes Administrator (CKA)

Experience contributing to Kubernetes ecosystem projects

Experience building multi-tenant platform infrastructure

Experience working in high-security and/or SOC 2/ISO27001 compliant environments

Experience with chaos engineering practices or implementation

Commitment to Equal Opportunity
Chainlink Labs is an equal opportunity employer. All qualified applicants will receive equal consideration for employment in compliance with applicable laws, regulations, or ordinances. If you need assistance or accommodation due to a disability or special need when applying for a role or in our recruitment process, please contact us via this form.

Global Data Privacy Notice for Job Candidates and Applicants
Information collected and processed as part of your Chainlink Labs Careers profile, and any job applications you choose to submit, is subject to our Recruiting Privacy Policy. By submitting your application, you are agreeing to our use and processing of your data as required.

#J-18808-Ljbffr

📌 Senior Site Reliability Engineer, Node Platform (British Columbia)
🏢 Framework Ventures
📍 British Columbia

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability engineer, node platform (british columbia) / british columbia

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability engineer, node platform (british columbia) / british columbia