Empower cutting-edge AI solutions at Baseten as an AI Model Engineer. Focus on the development and optimization of APIs essential for high-performance model inference.Join Baseten's Model Performance team, where your work will ensure fast, reliable, and cost-productive deployment of AI models. Ideal for candidates skilled in distributed systems and backend infrastructure, this role is pivotal in shaping how developers interact with AI technologies.Key Responsibilities:Architect and manage Model APIs with advanced inference featuresOptimize performance of TensorRT-LLM and CUDA kernelsDrive production-level performance enhancement initiativesDevelop comprehensive benchmarking frameworks for AI modelsFoster collaboration to deliver robust model serving solutionsRequirements:3+ years in managing distributed systems or APIsTrack record of low-latency backend service developmentProfound knowledge of profiling, tracing, and SLO managementExpertise in troubleshooting complex system interactionsStrong writing skills for effective design documentationBecome a key player in advancing AI model performance at Baseten.#J-18808-Ljbffr
📌 Ai Model Engineer At Baseten (Montreal)
🏢 Baseten
📍 Montreal
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.