Hiring: AWS Trainium / AI Infrastructure Engineer
Location: Canada — Remote
:
We are looking for an experienced AI Infrastructure Engineer with strong expertise in AWS Cloud, large-scale AI/ML platforms, and distributed model training.
The ideal candidate will have hands-on experience with AWS Trainium/Inferentia, AWS Neuron SDK, SageMaker, EKS/Kubernetes, and scalable AI infrastructure. Experience working with AI accelerators and optimizing large-scale ML workloads is highly desirable.
Key Skills:
- Hands-on experience with AWS Trainium, Inferentia, EC2 Trn instances, and AWS Neuron SDK for AI model training and inference.
- Strong understanding of LLMs, Generative AI, distributed training, and AI/ML infrastructure.
- Experience with Kubernetes (EKS), Docker, Python, PyTorch, and cloud-native architectures.
- Experience in AI performance optimization, benchmarking, and migration from GPU-based environments to AWS AI accelerators.
- Strong understanding of scalable and production-grade AI infrastructure on AWS.
Experience Required:
- 5+ years of experience in Cloud, Data, or AI Infrastructure.
- Experience supporting large-scale AI/ML training workloads in AWS environments.
- Hands-on experience with Trainium or Inferentia is highly preferred.
- Experience with distributed ML workloads, performance tuning, and cloud infrastructure optimization is a solid plus.
Ideal Candidate:
Candidates who have worked on LLM training/inference, GPU-to-Trainium migration, AWS Neuron optimization, or large-scale distributed AI platforms are strongly encouraged to apply.
? Interested candidates can share their updated resume for consideration.
[email protected] /
[email protected]
Contact No - (phone hidden)
📌 AWS Trainium / AI Infrastructure Engineer (Canada)
🏢 Aptino
📍 Canada