A technology company in Toronto seeks a Senior Site Reliability Engineer to manage and optimize HPC cluster operations, deploying infrastructure-as-code solutions. Candidates should have over 5 years of SRE or HPC operations experience, proficiency in Linux systems, and experience with Kubernetes and Ceph deployments. This role offers a market-competitive salary range of CAD 150,000 to 250,000 annually for skilled problem-solvers with a passion for innovation.#J-18808-Ljbffr
📌 Senior Sre: Ai/Ml Hpc Infrastructure & Gpu Cluster - C$150,000 - C$250,000 A Year (Toronto)
🏢 Boson AI
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.