25 Sep
|
DRW Holdings
|
Montreal
25 Sep
DRW Holdings
Montreal
Join DRW as a DevOps HPC Specialist and work on optimizing GPU infrastructures for cutting-edge AI and machine learning systems. Be part of a collaborative team that thrives on innovation.As an HPC Specialist at DRW, based in Chicago, you will focus on deploying and maintaining GPU infrastructures designed to support advanced AI and ML workloads. You'll use your extensive experience in infrastructure engineering to optimize performance and scalability, ensuring productive operation across complex systems. Collaboration with machine learning engineers will be key to improving model performance.Key Responsibilities:Optimize GPU infrastructure for LLM inferenceManage multi-GPU deployments in KubernetesConfigure and manage network componentsImplement storage solutions for model efficiencyTroubleshoot technical performance issuesRequirements:Bachelor's or Master's in Computer ScienceOver 5 years of experience in infrastructure rolesStrong skills in GPU management and optimizationProficient with Ansible or similar toolsExcellent understanding of distributed systemsLeverage your expertise in high-performance computing to deliver cutting-edge solutions at DRW.
📌 Devops Hpc Specialist At Drw (Montreal)
🏢 DRW Holdings
📍 Montreal