Senior Staff Reliability Engineer

ServiceNow - Santa Clara, CA

Hiring: Senior Staff Reliability Engineer Company: ServiceNow Location: Santa Clara, CA Job Posted Time: 2026-09-16 19:57:12 Employment Type: Contract / Remote Target Skills & Keywords : AWS, Agile, Ansible, ArgoCD, Azure, CI/CD, Configuration Management, Cypress, EKS, Feature Flags, GCP, GitLab CI, GitOps, Helm, Infrastructure as Code, Istio, Java, Kubernetes, Linkerd, OpenTelemetry, Playwright, Prometheus, Python, REST, Ruby, Selenium, ServiceNow, Systems Design, Terraform, TestNG, pytest About the job Experience: •12+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering with a Bachelor's degree; or 8 years and a Master's degree; or a PhD with 5 years experience; or equivalent experience. Required Skills: •Join us to put AI to work for people. •Join us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation, observability, and AI-driven operations. •What You Get To Do In This Role •Design, build, and operate cloud-native engineering platforms for software validation, release validation, and production readiness •Design and maintain production-like release and test ServiceNow environments that improve release confidence and deployment readiness. •Build and integrate automated test pipelines, observability, reliability signals, deployment intelligence, and quality gates into CI/CD workflows. •Develop automation solutions that improve engineering productivity, streamline operations, and reduce manual toil through shift-left engineering practices. •Build reusable frameworks, self-service engineering environments, test data management, mock services, and developer productivity tooling. Qualifications: •Applied hands-on capability in Kubernetes across cluster operations, networking, storage, security, autoscaling, and multi-cluster environments. •Strong software engineering skills with hands-on experience designing, developing, testing, and debugging applications using Python, Go, Java, or Ruby. •In-depth knowledge of observability, monitoring, SLI/SLOs, incident management, and production operations for distributed systems. •Demonstrated ability to solve complex technical problems, drive projects independently, and collaborate effectively across engineering teams. •Thrives in fast-paced, ambiguous environments with a strong ownership mindset, bias for action, and a passion for continuous learning and automation. •Low ego, intellectually curious, and an effective collaborator who enjoys partnering with globally distributed teams to deliver reliable engineering solutions. Compensation: •$190,900 - $334,100 / year •Flexible work environment (work from home / hybrid options) •Plus equity (when applicable), variable/incentive compensation and benefits Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!