Software Engineer, Machine Learning Infrastructure at Stripe.About the role Join the team responsible for the foundational systems that power machine learning across Stripe. You will build the infrastructure that supports the entire lifecycle of AI models, from initial data exploration to production deployment and LLM integration.Key factsLocation: Toronto, CanadaEngagement: Full-timeTeam: Machine Learning InfrastructureWhat you'll doArchitect and maintain secure, reliable services for model training, experimentation, and LLM applications across global regions.Develop libraries and internal tools that help engineers move models from research environments into production.Collaborate with product and data science departments to increase developer velocity.Manage technical projects that span multiple systems and operational requirements.Requirements2+ years of qualified software engineering experience focusing on distributed systems and service oriented architecture.Full lifecycle development experience,
including user requirements, design, implementation, testing, and production operations.Practical experience with MLOps, production ML platforms, or LLM application development.Background in managing high availability, low latency production systems.Proven ability to partner with cross-functional teams to achieve business goals.Ability to balance technical idealism with pragmatic execution.Nice to haveExperience developing and deploying production AI agents.Familiarity with LLM frameworks and large language models.History of training and shipping ML models to address specific business challenges.Skills & toolsDistributed systemsService oriented architectureMLOpsLLM frameworksMachine learning model training and servingHigh availability systems designPractical notesYou will coordinate with teams based across the US and Canada.#J-18808-Ljbffr