Advance your career with AMD as a Senior AI Systems Engineer. Take charge of GPU model execution, optimizing large-scale AI training and inference across cutting-edge hardware.
As part of the AMD AI Group, you will oversee the end-to-end execution stack on AMD Instinct GPUs, supporting transformative AI/ML technologies. This role requires proficiency in developing GPU kernels, orchestrating training infrastructures, and managing high-performance inference systems. Your expertise will drive innovation in training optimization, monitoring infrastructure, and inference serving frameworks, making a significant impact on the future of supercomputing.
Key Responsibilities:
• Enable large-scale AI model training on AMD Instinct GPU clusters
• Build training infrastructure like job orchestration and data pipelines
• Optimize communication patterns for multi-GPU training workloads
• Develop monitoring and compliance infrastructure for training clusters
• Drive inference optimizations and frameworks on AMD GPUs
Requirements:
• Robust experience in AI/ML infrastructure
• Proven record with frontier models on AMD hardware
• Knowledge of AMD Instinct architectures and RCCL
• Experience designing GPU validation frameworks
• Degree in Computer Science or related field
Leverage your technical expertise in AI and GPU technologies to lead at AMD and shape the future of computing.
#J-18808-Ljbffr
📌 Senior AI Systems Engineer at AMD (Ontario)
🏢 Advanced Micro Devices
📍 Ontario
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.