Who you are
- 5+ years of experience in Software Engineer, DevOps, SRE, and/or Platform Engineering roles
- Experience improving developer productivity through platform tooling, infrastructure automation, self-service capabilities, or internal developer platforms. Owns projects from idea through to production
- AI-first: Professional experience safely applying AI to infrastructure and platform engineering projects. Opinionated about safe AI usage around high-impact production workloads
- Strong cloud technology fundamentals: Familiarity with AWS, Kubernetes, IAC, and modern CI/CD tooling
- Senior-level experience administratively managing core service infrastructure: Servers, databases, caches, message queues, and data networks
- Demonstrated commitment to operational excellence including system reliability, on-call participation, incident response, and continuous improvement practices
- Comfortable working in a high-speed scale-up setting with broad scope, high autonomy, and rapid asynchronous remote communications
What the job involves
- Build and maintain our internal developer platforms: self-serve infrastructure, paved paths, and AI-first automation. Automation code to be authored primarily in Python and JS/TS
- Maintain and extend our cloud infrastructure: Particularly Kubernetes (EKS), AWS, Postgres, Redis, Kafka, and Elasticsearch. Leverage tools like Terraform and Crossplane to keep our infrastructure declarative
- Support other engineers as a go-to technical expert, and bake tribal knowledge into tooling and docs so the platform, not a single person, is the source of truth
- Contribute as a technical mentor: Directly mentor interns. Level up other engineers through Lunch & Learns, guilds, documentation, and general knowledge-sharing
- Contribute strategically during objective setting, retrospectives, and cross-team discussions, and help shape the infrastructure roadmap
- Help maintain the availability, reliability,
and security of our production infrastructure by participating in our on-call rotation, including monitoring, incident response, and resolving issues outside of standard business hours
- Note: On-call is a core part of this role. Engineers take turns in a weekly on-call rotation and are compensated for each rotation per the company’s on-call compensation policy
- Services are authored as K8s-based microservices powered by Python FastAPI backends and Typescript React front-end code
- Apps are built upon Postgres, Redis, Kafka, Elasticsearch, and Snowflake
- Our infrastructure runs on AWS and leverages hosted infrastructure technologies such as EKS, MSK, Elasticache, and RDS. We use Istio for service mesh, Helm, TF and Crossplane for IAC, and Kyverno for policy-as-code enforcement
- Trunk-based development and deploy-on merge through CI/CD is orchestrated with GitLab and ArgoCD, leveraging AI within the pipeline to deliver upwards of 1000 production deployments/month
- We invest heavily in reliability and observability using Datadog for monitoring and automated alerting, and use Amplitude for client-side analytics and experimentation
- Developer productivity and DevX are improved by this team through our CDE, Coder, DX360 surveys, and significant in-house productivity tooling
- As an AI-First engineering organization, we leverage modern AI development workflows using tools such as Cursor, Claude, and internally built AI agents to accelerate implementation and iteration across our codebases. You'll contribute directly to our AI infrastructure in this role
Benefits
- Competitive base salary
- Yearly learning & development allowance
- Bi-annual performance &
compensation reviews
- Generous equity options
- RRSP & 401k employee contribution plan
- Health, dental, & vision insurance on day one
- Parental leave programs
- Wellness reimbursement
- Supplemental Life Insurance
- Unlimited PTO (yes, really!)
- Recharge days throughout the year
- Permanent remote flexibility
- Work-from-home allowance
- Office in Toronto
📌 Senior Software Engineer (Canada)
🏢 super
📍 Canada