Lead Site Reliability Engineer

Glean - Mountain View, CA

Hiring: Lead Site Reliability Engineer Company: Glean Location: Mountain View, CA Job Posted Time: 2026-09-10 12:58:43 Employment Type: Hybrid Target Skills & Keywords : AWS, Azure, Docker, Embedded Systems, GCP, GDPR, GitHub, Infrastructure as Code, Kubernetes, LLM, SaaS, ServiceNow, Systems Design, Terraform, Zendesk, Zoom About the job Experience: •8+ years of experience in a senior-level role within Site Reliability Engineering or similar role, particularly in managing cloud-based services and infrastructure. •5+ years of experience with software development in one or more programming languages. •3+ years of experience managing people or teams, leading projects, and designing, analyzing, and troubleshooting distributed systems running in Cloud. Required Skills: •Ensure High Availability: Implement and maintain resilient cloud architectures, monitor system performance, and proactively identify and resolve potential bottlenecks or points of failure. •Incident Management: Participate in primary oncall rotation; cultivate technical curiosity and growth mindset, and a blameless postmortem culture within the team. Continuously optimize the on-call process for sustainability and efficiency. •Automation and Tooling: Develop and maintain automation scripts, tools, and processes to streamline system deployment, monitoring, and management tasks. Your contributions will be vital in efficiently scaling cloud operations. •Performance Optimization: Optimize cloud infrastructure and applications for performance, scalability, and cost-effectiveness. •Security and Compliance: Collaborate with security engineers to implement best practices and ensure compliance with security standards and policies. •Monitoring and Alerting: Design and configure advanced monitoring systems to gain insights into system behavior, set up alerts, and respond proactively to potential issues. Create and maintain comprehensive dashboards and playbooks for production on-call. •Software Development Consultation: Engage actively in the entire software development lifecycle. Participate in system design reviews and provide valuable SRE insights during launch reviews, influencing and enhancing system architecture. Qualifications: •Bachelor’s degree in Computer Science, a related field, or equivalent practical experience. •Strong knowledge of cloud platforms such as Google Cloud Platform, AWS, or Azure. •Practical experience with containerization technologies, including Docker and Kubernetes. Familiarity with infrastructure as code tools like Terraform is essential. •Solid understanding of networking, security principles, and best SRE and security practices. •Proficiency in using monitoring and alerting tools to detect and respond to potential issues effectively •This role is hybrid (4 days a week in our Mountain View Office) Compensation: •$200,000 - $260,000 / year •Competitive benefits and rewards package •Certain roles may be eligible for variable compensation, equity, and benefits Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!