21 Aug
|
Veeda AI
|
Ontario
Veeda AI in Toronto is seeking a Member of Technical Staff - ML Performance to own step time and model throughput for multi-node video world model training, optimize tensor usage, and implement advanced parallelism.
You will write and tune CUDA and Triton kernels, improve precision and numerical stability, and build fault diagnostics and elastic checkpointing to minimize downtime during interruptions.
#J-18808-Ljbffr
📌 Senior ML Performance Engineer - Distributed Training (Ontario)
🏢 Veeda AI
📍 Ontario