Who you are
- An experienced technical leader with substantial backend architecture and developer tooling experience to work alongside our team, develop and deliver on a technical roadmap for how code gets written and shipped at Faire, and engage with engineering teams organization-wide
- Extensive experience designing, building, and running distributed backend systems in production, including finding and permanently fixing problems at scale
- Experience owning the design of developer tools, CI/CD systems, or internal platforms that other engineering teams build on and hundreds of engineers use, and making design decisions that span several teams
- Hands-on experience building on large language models and coding agents: running agents, giving them tools, checking their output, and tracking what they do and what they cost. You use AI coding tools every day
- Experience running services with uptime targets: defining SLAs, building monitoring and alerting, and owning incidents for systems other developers depend on
- Good judgment on performance and cost tradeoffs, with a record of making a system cheaper to run without making it worse
- A record of mentoring engineers and raising the level of the teams you have worked on, both on your immediate team and across the organization
- Robust communication skills; ability to bring disparate groups of people to a shared understanding
- A bachelor's degree in Computer Science/Software Engineering or equivalent industry experience
- An advanced understanding of Kotlin; working knowledge of Python
- Experience running workloads across more than one cloud provider
What the job involves
- Our Engineering organization owns the software that makes our marketplace work
- Our Engineering Platform group enables product engineering teams to build and operate software with unmatched speed and quality
- The Code Authoring team owns the developer environment at Faire: the AI coding tools engineers use every day, remote development environments, and the sandboxes our end-to-end tests run in
- We care about developer experience, the quality of code written by people and agents, the reliability of the tools engineers depend on, and the cost of running those tools
- ---The team focuses on the following areas and capabilities:
- AI code authoring harnesses
- Cloud environments for running agents on Claude and Cursor
- Validation of agent-written code: CI checks, pre-push checks, and automated review
- Agent memory and context management
- Model selection and LLM cost per pull request
- Remote development environments
- Sandboxes for end-to-end testing and agent runs
- Pre-push validation speed
- Tracing and cost visibility for every agent run
- Uptime and SLAs for developer tooling
- Developer cloud infrastructure cost across AWS and OCI
- Integrations with Linear and GitHub
- The team also evaluates AI coding tools and owns our relationships with those vendors
- Own the architecture of the systems that run coding agents against our codebase: the service that turns tickets into pull requests, the sandboxes and remote development environments it runs in, and the integrations with Linear, GitHub, Claude, and Cursor. You decide how these pieces fit together and hold the long-term design
- Own the developer workflow end to end: how an engineer or an agent gets from a ticket to a merged, tested change. Know where it is slow or fragile, and change the tooling to fix it
- Run these systems as production infrastructure, with clear ownership, monitoring, uptime targets, and accountability when they fail. Developers across Faire depend on them every day
- Own what the platform costs to run:
know what each agent run and each development environment costs, and keep that spend proportional to what it delivers
- Set the standard for how teams build: design reviews, testing, operational readiness, and the quality bar for code written by people and by agents
- Represent the pod to the rest of engineering: partner with other teams on the workflows they own, take part in planning and prioritization, and shape the roadmap around what the tooling can and cannot yet do
- Mentor engineers on the pod and across the organization
- ---Technologies we use & teach---
- (Experience in similar technologies is welcomed)
- Backend: Kotlin, Java, JUnit, Hibernate, Guice, Jersey, MySQL, CockroachDB, AWS
- Frontend: React, TypeScript/JavaScript
- Infrastructure & Data: Python, HTTP, JSON, Protocol Buffers, experimentation frameworks, analytics tooling
Benefits
- Comprehensive healthcare: Including Health, Dental, Vision and Disability for all our locations.
- Time off: Paid time off, holidays and company-wide "Faire Fundays".
- Parental leave: Generous parental and family leave, as well as fertility support benefits.
- Productivity support: Monthly stipends to help cover work from home connectivity needs.
- Annual learning grant: For personal and professional development, as well as unlimited access to training courses through LinkedIn Learning.
- Thoughtfully designed spaces: All of our offices have been designed with local in mind – from our architects to our coffee blends.
- Fitness and well-being benefits: Including monthly credit towards your wellness-related programmes.
- Mental health benefits: Including free access to Modern Health therapists and resources.
- Charitable matching: Faire will match up to £250 of your charity donations, every year.
- Career planning: Whether you want to grow as a leader, hone your craft or explore a new discipline, our career framework allows space for you to explore.
📌 Staff Software Engineer (Toronto)
🏢 Faire
📍 Toronto