Site Reliability Engineer (US - Remote)
AXON Networks - Irvine, CA
Hiring: Site Reliability Engineer (US - Remote) Company: AXON Networks Location: Irvine, CA Job Posted Time: 2026-09-03 10:46:35 Employment Type: Remote Target Skills & Keywords : Apache, Bash, CI/CD, DNS, Digital Transformation, Firmware, GCP, Git, GitOps, Grafana, Helm, Infrastructure as Code, Java, Kafka, Kubernetes, Linux, Load Balancing, Microservices, OpenTelemetry, Oracle Cloud, Prometheus, Pulsar, Python, Systems Engineering, TCP/IP, TLS, Terraform About the job Experience: •5+ years of experience in site reliability engineering, production engineering, DevOps, cloud infrastructure, systems engineering or a closely related role. Required Skills: •Combine software engineering with hands-on NOC operations to make the complete cloud-to-device service path observable, supportable and resilient at fleet scale. •Help establish practical SRE capabilities inside the NOC while partnering closely with Support, Operations, cloud and DevOps Engineering. •Participate in a sustainable on-call rotation and improve the NOC’s ability to diagnose customer-impacting issues. Qualifications: •Strong software or automation skills in Python, Go, Java, Bash or a comparable language, with experience producing maintainable operational code. •Hands-on experience operating distributed production systems in a public cloud environment and troubleshooting across application, infrastructure, network and device-integration layers. •Strong Linux, containers and Kubernetes fundamentals, including deployment behavior, resource management, networking and failure diagnosis. •Strong troubleshooting & debugging skills in Kubernetes platforms. •Operational familiarity with Prometheus, Grafana, OpenTelemetry or equivalent observability ecosystems. •Operational familiarity with Apache Pulsar or similar distributed messaging and streaming platforms handling requests from millions of devices. •Solid functional working knowledge of SLOs, error budgets, capacity planning, resilience engineering, change safety and blameless incident learning. •Strong networking knowledge, including TCP/IP, DNS, DHCP, TLS, routing, NAT, load balancing and systematic packet- or session-level troubleshooting. •Clear communication, disciplined documentation and the ability to collaborate across NOC, cloud, DevOps, firmware and service-provider teams. •Bachelor’s degree in computer science, engineering or equivalent practical experience. Compensation: •$160,000 - $200,000 / year •Flexible work environment (work from home / hybrid options) Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!