18 Sep
|
Stream
|
Toronto
Who you are
- 5+ years in infrastructure, platform, DevOps or SRE engineering, with clear depth in infrastructure over application development
- A software engineering background. You have built systems, not only configured them. Production coding experience in Go or Python. Scripting-only backgrounds are not a fit
- Kubernetes at meaningful production scale, past operations: you have driven cluster strategy, designed workloads, or led a migration, and you have tuned what came out the other side for cost and efficiency
- Cloud cost or efficiency optimisation you personally led on AWS or GCP, with an outcome you can put a number on. FinOps practice is a plus
- Direct experience running high-scale, high-load production systems
- Strong cloud fundamentals across networking, compute, storage and IAM, and the habit of asking why a system behaves the way it does instead of accepting the default
- Comfortable in a small team: leading a project and reviewing a PR in the same week
- AI tooling already in your engineering workflow. Applied use, not familiarity
- Both AWS and GCP, and migration experience between providers
- PostgreSQL at scale: sharding, replication strategy, partitioning tradeoffs, ideally self-hosted
- Real-time systems: WebSockets, WebRTC, streaming or other persistent-connection workloads
- The wider stack: CockroachDB, Redis, Terraform, and a Prometheus-based observability stack
- An API-first or infrastructure company at scaleup stage
- Open source contributions to infrastructure or platform tooling
- Writing or talks on cloud, platform or distributed systems
- Formal FinOps practice, or owning cloud commitment and reservation strategy
- Work on developer-facing API or SDK products
What the job involves
- We are hiring a Senior Software Engineer to help rebuild the platform underneath Stream
- Over the next year the infrastructure team is moving from AWS to GCP, moving onto Kubernetes, and relocating 35 to 40 Postgres shards off managed RDS to self-hosted, while the platform keeps serving billions of API requests a month
- You will own parts of that outright
- This is a small, senior team without the support structures of a large organisation
- You will write code most of the time and make infrastructure calls on your own
- Success looks like systems that scale predictably under load, cloud spend that falls per unit of traffic, and migrations that land without incident
- Design, build and operate infrastructure for real-time systems carrying millions of concurrent connections and billions of monthly API requests
- Drive Kubernetes end to end: cluster architecture, workload design and the migration of existing services. You will be designing clusters, not operating someone else's
- Re-architect workloads as part of the AWS to GCP migration, for cost and performance rather than a lift and shift
- Own cloud cost and efficiency work: find the levers, measure them against real spend and utilisation data, and show what moved
- Write production Go and Python: internal services, platform tooling and automation that change how product and SDK engineers deploy,
observe and debug
- Lead post-migration tuning and capacity planning, closing the loop between the architecture you chose and what production actually does
- Work with backend, video and moderation engineers on system design, reliability targets and tradeoffs that cross service boundaries
- Take part in on-call, incident response and root cause analysis, and turn what you find into durable fixes
- ---Our stack---
- Go, gRPC, RocksDB, Python
- PostgreSQL, RabbitMQ
- GCP
- Grafana, Prometheus, ELK (Elasticsearch and Kibana)
- Jaeger and Tempo for distributed tracing, Datadog
- Redis, Memcached
- Claude Code, Cursor
Benefits
- Equipment & Resources: Computers, standing desks, headphones, and other specialized tools are available for you to be successful
- Paid Vacation: Generous time off on top of observing several national holidays
- Stock Options: You get a piece of ownership in Stream. Stock options are granted to all full-time employees
- Health Coverage: We offer comprehensive healthcare, dental, and vision plans, plus generous parental leave (US only)
- Outings and parties: We like to have fun together! Team outings and parties are regular events around here
- Team Lunch and Snacks: Stream provides in-office daily team lunches and a stocked kitchen with tasty snacks and craft coffee
- 401k: Plan for your retirement through the company offered 401k (US only) or the local pension scheme in Amsterdam, to which Stream will make a contribution
- Commuter Benefits: Employees receive free Office parking and transportation perks
- Monthly fitness stipend: We value the health and well being of our employees
📌 Senior Software Engineer (Toronto)
🏢 Stream
📍 Toronto