AI Platform Operations Manager
STACK Infrastructure - Denver Metropolitan Area
Hiring: AI Platform Operations Manager Company: STACK Infrastructure Location: Denver Metropolitan Area Job Posted Time: 2026-09-16 15:38:14 Target Skills & Keywords : Azure, Azure DevOps, Bash, CI/CD, Configuration Management, Cosmos DB, Databricks, Docker, Foundry, Git, GitHub Actions, Grafana, Infrastructure as Code, Kubernetes, LLM, MLOps, MLflow, Machine Learning, NetSuite, Node.js, PowerShell, Prometheus, Python, RAG, Root Cause Analysis, Terraform, Vault, Workday About the job Experience: •5+ years of hands-on experience in DevOps, site reliability engineering, or platform engineering roles with a strong delivery track record. Required Skills: •Infrastructure Automation & Infrastructure as Code •Build, maintain, and version infrastructure-as-code modules for Azure environments using Terraform, Bicep, or ARM, including compute, networking, storage, identity, and AI platform resources. •Automate provisioning of AI platform components — Azure AI Foundry, Azure OpenAI Service, Azure AI Search, Cosmos DB, ADLS Gen2, and Databricks — as reusable, parameterized deployment patterns. •Maintain environment parity across development, test, and production, including configuration management, drift detection, and remediation. •Implement and enforce tagging, naming, and resource organization standards that support governance, chargeback, and lifecycle management. •Automate routine platform operations — patching, certificate rotation, key and secret rotation, backup validation, and disaster recovery testing. •Design, build, and operate CI/CD pipelines in Azure DevOps or GitHub Actions for application code, infrastructure code, container images, and AI/agent deployments. •Implement automated build, test, security scanning, artifact management, and promotion gates across environments. Qualifications: •Bachelor's degree in Computer Science, Information Technology, Engineering, or related field, or equivalent practical experience. •Strong proficiency in Infrastructure as Code — Terraform, Bicep, or ARM — including module design, state management, and reusable patterns. •Proven experience building and operating CI/CD pipelines in Azure DevOps or GitHub Actions. •Applied hands-on capability in containerization and orchestration — Docker, Azure Kubernetes Service (AKS), and Azure Container Apps or equivalent. •Solid working knowledge of Azure core services — compute, networking (VNet, NSG, Private Endpoints), storage, identity (Entra ID), and Key Vault. •Strong scripting and automation skills in Python, PowerShell, or Bash. •Solid functional working knowledge of Git-based workflows, code review practices, and artifact/registry management. •Demonstrated ability to troubleshoot production issues across infrastructure, network, and application layers. •Microsoft Certified: DevOps Engineer Expert (AZ-400), Azure Administrator (AZ-104), or Certified Kubernetes Administrator (CKA). •Operational familiarity with MLOps tooling and practices — Azure Machine Learning, MLflow, Databricks, or equivalent model lifecycle platforms. Compensation: •$128,260 - $146,017 / year Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!