Join us as an ML Engineer specializing in infrastructure dedicated to advancing machine learning technologies. This role focuses on building systems that facilitate high-throughput research on large language models. In this cutting-edge position, you will create, scale, and optimize the infrastructure that supports our pioneering research, helping address real-world challenges in ML.
Familiarity with cloud services and distributed systems techniques is crucial as you enhance our capabilities in handling complex ML tasks. Work closely with engineers to develop robust internal tooling and frameworks. Key Responsibilities:
Build and maintain infrastructure for ML post-training research
Design evaluation and benchmarking tools for quality assurance
Develop automated testing and deployment systems
Collaborate with researchers to fulfill infrastructure requirements Requirements:
Proficiency in PyTorch or JAX frameworks
Experience with cloud platforms like AWS or GCP
Hands-on knowledge of data engineering tools
Some familiarity with LLM training internals
Solid communication skills for collaborating with research teams Support groundbreaking ML research by enhancing our infrastructure and processes.
📌 Ml Engineer For Scalable Infrastructure Toronto
🏢 United States Digital Space
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.