EnCharge AI is seeking an experienced AI Research Engineer to optimize deep learning models for edge-to-cloud deployment. You will focus on model compression, quantization, and efficient inference techniques to boost AI workloads. Strong collaboration with hardware teams is essential.
Required expertise includes QAT/PTQ, mixed-precision strategies, and proficiency in PyTorch, CUDA, and C++. This role offers a challenging, cutting-edge setting with significant impact on performance and energy
📌 Edge AI Research Engineer: Model Compression & Inference (Canada)
🏢 United States Digital Space
📍 Canada
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.