Cerebras Systems is building a current generation of disaggregated AI inference systems. We are hiring a Software Engineer to evolve the ML API layer for a heterogeneous serving system across models and accelerator backends.
You will work across inference APIs, model integration, vLLM-based runtime, and Cerebras runtime components to deliver a consistent experience for developers and customers. This role is hands-on and focuses on production-quality APIs.
📌 Staff AI Inference Systems Engineer (Toronto)
🏢 Engg
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.