GPU Systems Engineer – HPC / Parallel Computing

Vast.ai - Los Angeles, CA

Hiring: GPU Systems Engineer – HPC / Parallel Computing Company: Vast.ai Location: Los Angeles, CA Job Posted Time: 2026-09-10 10:29:34 Employment Type: Full-time / On-site Target Skills & Keywords : C++, LLM, Linux, Python About the job Experience: •Expertise with at least one parallel framework (CUDA, HIP, SYCL, OpenCL, OpenACC, or similar) •Strong background in systems optimization and HPC performance tooling •Operational familiarity with distributed training/inference frameworks (bonus) •45 min - Quick dive into Vast, work history (virtual) •45 min - Systems and architectures (virtual) Required Skills: •We’re looking for a systems engineer with HPC or parallel programming experience to help scale AI inference. You’ll leverage your knowledge of high-performance systems to optimize GPU performance at the bleeding edge of AI. •On-site at either our SF or LA offices •Design and optimize GPU kernels and tensor libraries •Translate HPC techniques into scalable AI inference solutions •Evaluate emerging architectures and resource management approaches •Partner cross-functionally with technical leadership to improve GPU infrastructure efficiency Compensation: •$160,000 - $320,000 / year •Comprehensive health, dental, vision, and life insurance Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!