19 Aug
|
ClickHouse
|
Mississauga
19 Aug
ClickHouse
Mississauga
Senior Site Reliability Engineer- Remote
Location: Canada (remote)
Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies. With more than 3,000 customers and ARR that has grown over 250 percent year over year, ClickHouse leads the market in real-time analytics, data warehousing, observability, and AI workloads.
The company’s sustained, accelerating momentum was recently validated by a $400M Series D financing round. Customers include Capital One, Lovable, Decagon, Polymarket, and Airwallex, in addition to brands such as Meta, Cursor, Sony, and Tesla. We’re on a mission to transform how companies use data.
We are expanding our central Site Reliability Engineering team to provide reliable and secure services. You will be responsible for building and leading processes to ensure the reliability, availability, scalability, and performance of our cloud infrastructure. You will collaborate with Control Plane, Data Plane, Core, Security, Support and Operations teams to design and implement scalable, secure, highly available and fault-tolerant distributed systems.
You will own incident management and response, post-mortem analysis including blameless postmortems, and continuous improvement of our Cloud services. You will leverage software engineering to develop platforms and tools to optimize operational and engineering efficiencies of ClickHouse Cloud. This role offers the prospect to impact our elastic, high-performance ClickHouse Cloud at scale.
Collaborate with engineering teams to design and implement scalable, secure,
and highly available systems for ClickHouse.
Ensure infrastructure components in ClickHouse Cloud have monitoring and alerting to detect and resolve incidents.
Improve incident response processes and post-mortem analysis, including communicating with impacted customers through the support team.
Continuously improve reliability and performance of ClickHouse services.
Plan, enable, and drive Chaos initiatives across Engineering teams based on internal priorities.
Bachelor’s or Master’s degree in Computer Science or related field.
At least 8 years of experience in Site Reliability Engineering or related field.
Hands-on experience with Go and/or Python.
Strong knowledge of cloud platforms (AWS, Azure, GCP).
Excellent understanding of distributed databases and SQL;
Experience with automation and configuration management tools (Ansible, Terraform, Puppet).
Focus on efficiency, availability, scalability, and data governance.
These ranges reflect the minimum and maximum pay at posting and may be adjusted in the future. An individual’s placement within the range depends on factors including education, qualifications, certifications, experience, skills, location, performance, and business needs.
Flexible work environment - Remote-friendly, 20 countries in operation.
Equity - Flexible time off in the US, generous entitlement elsewhere.
A $500 Home office setup - for remote employees.
For government reporting, we invite candidates to respond to the voluntary self-identification survey. If you have questions, visit the OFCCP website. #
📌 Senior Site Reliability Engineer-Remote (Mississauga)
🏢 ClickHouse
📍 Mississauga