EnCharge AI is seeking an experienced AI Research Engineer to optimize deep learning models for edge-to-cloud deployment. You will focus on model compression, quantization, and effective inference techniques to boost AI workloads. Robust collaboration with hardware teams is essential.
Required expertise includes QAT/PTQ, mixed-precision strategies, and proficiency in PyTorch, CUDA, and C++. This role offers a challenging, cutting-edge workplace with significant impact on performance and energy
📌 Edge Ai Research Engineer: Model Compression & Inference Winnipeg
🏢 United States Digital Space
📍 Winnipeg
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.