Advance your career with AMD as a Senior AI Systems Engineer. Take charge of GPU model execution, optimizing large-scale AI training and inference across cutting-edge hardware. As part of the AMD AI Group, you will oversee the end-to-end execution stack on AMD Instinct GPUs, supporting transformative AI/ML technologies.
This role requires proficiency in developing GPU kernels, orchestrating training infrastructures, and managing high-performance inference systems. Your expertise will drive innovation in training optimization, monitoring infrastructure, and inference serving frameworks, making a significant impact on the future of supercomputing. Key Responsibilities:
- Enable large-scale AI model training on AMD Instinct GPU clusters
- Build training infrastructure like job orchestration and data pipelines
- Optimize communication patterns for multi-GPU training workloads
- Develop monitoring and compliance infrastructure for training clusters
- Drive inference optimizations and frameworks on AMD GPUs
Requirements:
- Solid experience in AI/ML infrastructure
- Proven record with frontier models on AMD hardware
- Knowledge of AMD Instinct architectures and RCCL
- Experience designing GPU validation frameworks
- Degree in Computer Science or related field
Leverage your technical expertise in AI and GPU technologies to lead at AMD and shape the future of computing.
📌 Senior AI Systems Engineer at AMD (Markham)
🏢 Advanced Micro Devices
📍 Markham
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.