About the Site Reliability Engineering Team
Thumbtack's Site Reliability Engineering team focuses on creating and maintaining a reliable, secure, and scalable platform vital for a seamless user experience. As a key contributor, you will design and support resilient systems, prioritizing high performance, availability, and throughput, with a focus on minimizing service disruptions, downtime, and latency. SRE impacts Thumbtack’s ecosystem across the entire stack, from linux systems to applications that drive the customer experience.
Our work is high leverage impacting how Engineering, Applied Science, and many other teams deliver, run, and observe systems.
The Site
Reliability team is responsible for a broad set of technologies and systems with expectations to collaborate across the business. We are expected to develop and enhance existing capabilities while ensuring scalability, reliability and resiliency of infrastructure and software. You’ll work with engineering teams ranging from product development, developer experience, and backend infrastructure to collaboratively build Thumbtack’s ecosystem of platform services that have the right impact at the right time.
Thumbtack values its cross functional team-oriented culture, and you’d be positioned to contribute to the future direction and success of the engineering platform that serves as the engine of our applications. Design, create, and maintain software and systems to improve the availability, scalability,
and efficiency of Thumbtack's services Set the architectural direction of infrastructure and platform services while supporting the engineering organization Design and implement tools and processes used for deployment, change, service, and infrastructure management Contribute to the evolution and performance of capabilities we provide to engineering as a platform organization Capacity planning and demand forecasting, anticipating performance bottlenecks Extensive fluency in AWS and Linux Ability to effectively read, write, and debug code in programming languages like but not limited to: Python, Go, PHP, Javascript Expertise in designing, analyzing, and troubleshooting large-scale distributed systems across web technologies like: Demonstrable knowledge of instrumenting, operating, and observing a distributed system of microservices in a production cloud environment Ability to communicate clearly and effectively to cross functional partners of various technical levels Thumbtack embraces diversity. We are proud to be an equal opportunity workplace and do not discriminate on the basis of sex, race, color, age, pregnancy, sexual orientation, gender identity or expression, religion, national origin, ancestry, citizenship, marital status, military or veteran status, genetic information, disability status, or any other characteristic protected by federal, provincial, state, or local law.
Thumbtack is committed to working with and providing reasonable accommodation to individuals with disabilities. If you would like to request a reasonable accommodation for a medical condition or disability during any part of the application process, please contact:
[email protected] you are a California resident, please review information regarding your rights under California privacy laws contained in Thumbtack’s Privacy policy available at #
📌 Senior Site Reliability Engineer - Software Engineering (Toronto)
🏢 Thumbtack
📍 Toronto