Site Reliability Engineer

Apex - Los Angeles, CA

Hiring: Site Reliability Engineer Company: Apex Location: Los Angeles, CA Job Posted Time: 2026-09-12 11:29:58 Employment Type: Full-time Target Skills & Keywords : Apex, ArgoCD, CI/CD, GitOps, Grafana, Infrastructure as Code, Kubernetes, Prometheus, Python, REST, Terraform About the job Experience: •5+ years of experience in site reliability, infrastructure, or network engineering, with a meaningful portion in aerospace, defense, satellite, or another mission-critical domain. Required Skills: •, setting the standards that the rest of the team builds on. •Architect and build ground and site network infrastructure spanning commercial and secured or classified environments through the platform layer. •Design, deploy, and scale highly available Kubernetes clusters that support workloads across multiple security and classification levels. •Own the security posture of ground infrastructure, defining and enforcing controls, hardening, and compliance for sensitive government and commercial programs. •Build and maintain infrastructure as code using Terraform, Terragrunt, and similar tooling so that environments are repeatable, reviewable, and fast to stand up. •Establish observability across the ground network, including monitoring, metrics, logging, and alerting, so issues are caught before they reach a mission. •Design and run CI/CD pipelines and automation that streamline deployments and reduce manual, error-prone work. •Act as a primary site reliability engineer for ground systems, driving uptime, incident response, and reliability standards on programs where downtime is not an option. Qualifications: •Applicants must be U.S. persons as defined by U.S. export-control law. •Hands-on experience architecting and operating ground network or large-scale production network infrastructure. •Deep expertise with Kubernetes and containers, including building and scaling high availability clusters in production. •Strong networking fundamentals across routing, switching, segmentation, and secure network design. •Proven experience with infrastructure as code with CI/CD, automation, and observability tooling such as Prometheus, Grafana, or similar. •Solid functional working knowledge of security and compliance for regulated or classified environments, and the judgment to design a defensible security posture. •Proficiency scripting and building tooling in Python, Go, or a comparable language. •Active Top Secret or Top Secret/SCI clearance •Prior experience supporting classified or government space programs or standing up infrastructure across multiple classification levels. •Operational familiarity with GitOps workflows and tools such as ArgoCD, and with modern observability stacks including Mimir. Compensation: •$155,000 - $195,000 / year Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!