As an AI-first company, Nuro’s ML engineers run a large number of distributed training jobs every day on a variety of computing resources. To support training at scale, we need a robust training infrastructure so that our ML engineers can train as efficiently as possible. One of the goals of our training infrastructure is to provide ML engineers access to accelerator resources, particularly GPUs and TPUs.