Site Reliability Engineer
Nextpoint - Chicago, IL
Hiring: Site Reliability Engineer Company: Nextpoint Location: Chicago, IL Job Posted Time: 2026-09-15 10:46:41 Target Skills & Keywords : AWS, CDK, CI/CD, CloudFormation, CloudWatch, Configuration Management, Datadog, Docker, EC2, ECS, Encryption, Grafana, HIPAA, IAM, Kubernetes, Lambda, Prometheus, Python, RDS, S3, SOC 2, SaaS, Terraform, VPC About the job Experience: •5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles, ideally in a B2B SaaS environment Required Skills: •This is a hands-on role for someone who wants to be the first line of response for production infrastructure at a company where reliability is a customer-trust issue, not just an engineering metric. •Maintain and extend existing infrastructure-as-code (Terraform/CloudFormation/CDK) following established patterns and standards •Support and operate CI/CD pipelines; implement improvements as directed •Monitor cloud cost trends and flag optimization opportunities for review •Monitor uptime, latency, and performance SLOs/SLIs for production systems supporting document upload, processing, review, and production workflows •Participate in the on-call rotation and serve as first responder for production incidents during US business hours •Triage, troubleshoot, and resolve incoming infrastructure requests and incidents; escalate and coordinate on complex root-cause work •Write clear post-incident reports and help investigate recurring incidents and cost overruns Qualifications: •Deep hands-on experience with AWS (EC2, S3, RDS, Lambda, VPC, IAM, CloudWatch, or equivalent services) •Solid functional working knowledge of infrastructure-as-code (Terraform, CloudFormation, or CDK) and configuration management •Proficiency in at least one scripting/programming language (Python, Go, or similar) for automation and tooling •Track record of operating CI/CD pipelines •Operational familiarity with security compliance frameworks (SOC 2, HIPAA, or similar) and encryption/access-control best practices •Exposure to AI/ML infrastructure (e.g., Amazon Bedrock, model-serving pipelines) is a plus, given our growing AI feature set •Bachelor's degree in Computer Science, Engineering, or related field (or equivalent experience) •Strong written and verbal communication skills, with the ability to write clear runbooks and explain technical tradeoffs to non-technical stakeholders •Comfortable being part of an on-call rotation Compensation: •$135,000 - $175,000 / year Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!