Principal Engineer, Efficient GenAI (Markham)

Principal Engineer, Efficient GenAI (Markham)

03 Oct
|
Advanced Micro Devices
|
Markham

03 Oct

Advanced Micro Devices

Markham

At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. The AI Models and Applications team at AMD is looking for a specialized Principal Engineer who is passionate about enabling innovative and efficient Generative AI training and inference at scale. You will be part of a core team of incredibly talented specialists and work on scaling training and inference for the latest Generative AI models.

You have a deep technical understanding and hands‑on experience with the latest Generative AI applications in at least one of the following areas: large language models (LLMs), 3D World and Action Models, or image/video generation models. You have experience training models at scale and are passionate about developing efficient approaches to enable distributed training and inference on AMD devices.

Exciting Opportunities: As a senior member of the team, you will be at the forefront of innovation, working with the latest Generative AI models and algorithms. You will have the opportunity to shape the future of AI model training and inference optimization across a variety of applications. Collaborate with like‑minded professionals and learn from the best in the field.

Cutting-Edge Technology: Work with state‑of‑the‑art Generative AI algorithms and software, enabling you to stay ahead of the curve and drive advancements in AI model training at scale and deployment.

Impactful Work: Your contributions will directly influence how cutting‑edge Generative AI models across the industry are efficiently trained at scale, as well as how inference solutions are deployed to serve millions of customers, making a significant difference across various industries and applications. Propose and apply innovative techniques to support both training and inference, including innovative transformer architectures, parallelism strategies for training on large clusters,



inference optimization techniques such as speculative decoding, and optimal KV‑caching strategies. Implement novel, efficient architectures for Generative AI models for training and inference and showcase the benefits on AMD platforms.

Work with open‑source frameworks and communities (e.g., PyTorch, JAX, vLLM, SGLang) to integrate AMD‑optimized models and libraries and publish training recipes. Collaborate with software and hardware teams to co‑optimize end‑to‑end performance on current and future AMD solutions. Solid technical expertise in Generative AI model training and inference, with familiarity working with deep learning frameworks such as PyTorch, JAX, vLLM, SGLang, and MuJoCo.

Strong technical expertise in algorithmic innovation for efficient Generative AI applications across both training and inference. Expertise and publications in one or more of the following preferred areas: effective model architectures, optimized training, innovative parallelism strategies, or low‑precision training.

Experience productizing Generative AI models and training foundation models at scale. Several years of experience in AI, deep learning, and related software development.

ACADEMIC CREDENTIALS: PhD or Master’s degree in Computer Science, Electrical Engineering, Mathematics, or a related field. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third‑party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

📌 Principal Engineer, Efficient GenAI (Markham)
🏢 Advanced Micro Devices
📍 Markham

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: principal engineer, efficient genai (markham) / markham

Subscribe to this job alert:

Get the latest job offers by email for: principal engineer, efficient genai (markham) / markham