Amazon's Annapurna Labs team is hiring engineers to design and optimize high-performance compute kernels for ML on Neuron. You will work across Neuron architecture generations, profiling and tuning for maximum throughput, and implement compiler optimizations while collaborating with customers to enable their models on AWS accelerators.
The role blends machine learning, high-performance computing, and distributed architectures in a startup-like setting, with mentorship and customer-facing