Senior Site Reliability Engineer
Cross River - United States
Hiring: Senior Site Reliability Engineer Company: Cross River Location: United States Job Posted Time: 2026-09-17 11:55:27 Target Skills & Keywords : .NET, AWS, Bash, Blockchain, CI/CD, CloudWatch, DNS, Datadog, Docker, ECS, ELK Stack, GitOps, Grafana, IaC, Infrastructure as Code, Istio, Linkerd, Linux, Load Balancing, New Relic, PowerShell, Prometheus, Python, Root Cause Analysis, SOC 2, Service Mesh, Terraform, Vault About the job Experience: •8+ years in SRE, DevOps, or Infrastructure Engineering roles •5+ years with AWS (preferred); experience with multi-cloud is a plus •Deep experience designing, building, and maintaining CI/CD pipelines and automation workflows •Strong Experience With Docker And Container Orchestration (ECS Preferred) •Proficiency with tools such as New Relic, ELK Stack, CloudWatch, Prometheus, Datadog, or Grafana Required Skills: •Define and enforce DevOps guardrails, standards, and best practices to ensure consistency, security, and compliance across engineering teams •Enable Engineering teams to design, implement, and maintain CI/CD pipelines best to enable fast, safe, and repeatable deployments across all environments •Co-Develop and maintain Infrastructure as Code (IaC) with Application teams using tools such as Terraform •Establish and govern deployment strategies including blue/green, canary, and rolling deployments •Build and maintain developer self-service tooling and internal platforms that accelerate delivery while maintaining governance •Champion a "shift-left" culture by embedding reliability, security, and observability practices early in the software development lifecycle •Help define, implement, and monitor Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets for critical services •Build and maintain comprehensive observability stacks including centralized logging, metrics, distributed tracing, and alerting using tools such as New Relic, ELK, Prometheus, and Grafana Qualifications: •Operational familiarity with SRE frameworks as outlined in Google's SRE handbook •Understanding of compliance and regulatory requirements in financial services •Financial industry / banking infrastructure experience is helpful, but not required •Crypto / blockchain infrastructure experience is helpful, but not required •Operational familiarity with cost optimization and FinOps practices in cloud environments •System uptime and availability targets consistently met or exceeded •Reduction in mean time to detect (MTTD) and mean time to resolve (MTTR) •Adoption and adherence to DevOps guardrails across engineering teams •Measurable reduction in operational toil through automation Compensation: •$160,000 - $200,000 / year Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!