Staff Software Engineer, ML Training and Inference Infrastructure
Rivian - Palo Alto, CA
Hiring: Staff Software Engineer, ML Training and Inference Infrastructure Company: Rivian Location: Palo Alto, CA Job Posted Time: 2026-09-03 10:28:31 Employment Type: Full-time Target Skills & Keywords : Deep Learning, Machine Learning, PyTorch, TensorRT, Triton About the job Required Skills: •Optimize the performance of Deep Learning training workload on NVIDIA GPU systems on a large scale •Optimize the latency of model inference and model pre- and post-processing on onboard systems •Design, train, and deploy large deep learning models that can leverage the vast amount of labeled and unlabeled data Qualifications: •PhD in CS/CE/EE, or equivalent, in industry experience •Knowledge of model training framework (e.g. PyTorch Lightning, ray, etc.) •In-depth knowledge of transformer architecture and ways to accelerate the training and inference of transformer models •A track record of profiling models and doing detective work to improve model training and inference speed •A track record of efficiently solving complex problems collaboratively on larger teams •Salary Range for California Based Applicants: $228,000.00 - $285,000.00 (actual compensation will be determined based on experience, location, and other factors permitted by law). •: Rivian provides robust medical/Rx, dental and vision insurance packages for full-time employees, their spouse or domestic partner, and children up to age 26. Coverage is effective on the first day of employment Compensation: •$228,000 - $285,000 / year Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!