Machine Learning Infrastructure Engineer, Model Inference
Abridge - San Francisco, CA
Hiring: Machine Learning Infrastructure Engineer, Model Inference Company: Abridge Location: San Francisco, CA Job Posted Time: 2026-09-16 15:37:18 Employment Type: Full-time Target Skills & Keywords : Ansible, Clinical Documentation, EMR, GitOps, Infrastructure as Code, Kubernetes, LLM, Machine Learning, PyTorch, TensorFlow, Terraform, Triton, vLLM About the job Experience: •5+ years of experience in building and deploying machine learning models in production environments. •5 years of employment. Required Skills: •Design, deploy and maintain scalable Kubernetes clusters for AI model inference and training •Develop, optimize, and maintain ML model serving infrastructure, ensuring high-performance and low-latency. •Partner cross-functionally with ML and product teams to scale backend infrastructure for AI-driven products, focusing on model deployment, throughput optimization, and compute efficiency. •Optimize compute-heavy workflows and enhance GPU utilization for ML workloads. •Build a robust model API orchestration system •Partner cross-functionally with leadership to define and implement strategies for scaling infrastructure as the company grows, ensuring long-term efficiency and performance. •5+ years of experience in building and deploying machine learning models in production environments. •Comprehensive expertise in container orchestration and distributed systems architecture Compensation: •$221,000 - $260,000 / year Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!