Create Alert
Email me similar jobs

Senior ML Performance Engineer - Distributed Training

Veeda AI in Toronto is seeking a Member of Technical Staff - ML Performance to own step time and model throughput for multi-node video world model training, optimize tensor usage, and implement advanced parallelism.

You will write and tune CUDA and Triton kernels, improve precision and numerical stability, and build fault diagnostics and elastic checkpointing to minimize downtime during interruptions.

#J-18808-Ljbffr
Similar jobs

Senior ML Performance Engineer - Distributed Training

Apply Now
Back to search page