Become a Distributed Systems Engineer at Cerebras Systems, enhancing extraordinary AI supercomputers. This role focuses on delivering exceptional software for large clusters and monitoring applications effectively.
As part of the Cluster engineering team, you'll automate processes, develop orchestration systems, and monitor system performance. Candidates should have a robust understanding of distributed computing, proficient development skills in GoLang and Python, and a grasp of Kubernetes. Your contributions will help unlock the speed and intelligence of AI applications, significantly transforming user experiences.
Key Responsibilities:
• Automate configurations of large-scale Cerebras clusters
• Enhance workflows for upgrades and vulnerability patches
• Design scheduling systems for multi-user environments
• Enable smooth operations for on-premise and cloud deployments
• Build systems for monitoring and failure detection
Requirements:
• Proven expertise in software architecture and design
• Experience in developing distributed cluster applications
• Deep understanding of Kubernetes, Prometheus, and Grafana
• Proficient in GoLang, Python, and bash
• Robust debugging skills across distributed systems
Transform the future of AI computing with your skills at Cerebras Systems.
#J-18808-Ljbffr
📌 Cerebras Systems Distributed Systems Engineer (Quebec City)
🏢 Cerebras Systems
📍 Quebec City
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.