Staff / Principal Platform Engineer

AppGate - New York, NY

Hiring: Staff / Principal Platform Engineer Company: AppGate Location: New York, NY Job Posted Time: 2026-09-16 18:38:58 Employment Type: Hybrid Target Skills & Keywords : AWS, CI/CD, Configuration Management, DNS, Docker, ELK Stack, Elasticsearch, Feature Store, Grafana, Helm, Infrastructure as Code, Kafka, Kubeflow, Kubernetes, MLOps, MLflow, OpenTelemetry, Prometheus, Python, SageMaker, TCP/IP, TLS, Terraform, VPN, Zero Trust About the job Experience: •Extensive platform, infrastructure or SRE engineering experience, with a track record of operating production systems at scale. Staff-level candidates typically bring 8+ years and Principal-level candidates 12+ years, though we hire on demonstrated impact •Strong command of infrastructure-as-code (Terraform or equivalent), CI/CD, containers and orchestration (Docker, Kubernetes), and cloud platforms (AWS) Required Skills: •As we expand our platform, we are standing up a new AI Platform & Infrastructure team: the engine room of AppGate's AI strategy. This team owns the infrastructure layer that every next-generation security capability is built on, from network observability to AI-driven threat detection and the secure operation of emerging Agentic AI systems. •You'll own the platform spanning APIs, cloud and self-managed solutions and AI/ML infrastructure, and you'll make it fast, reliable and observable at scale. This is a high-leverage, hands-on role for a senior engineer who sets technical direction and still ships. •Design, build and operate the cloud infrastructure, services and pipelines that AppGate's AI and cloud products run on. Strong experience with self-managed technologies (kafka, elasticsearch) and Kubernetes are a must •Terraform and Helm for cloud provisioning, service deployment and configuration management •Instrument APIs, cloud services and AI/ML infrastructure with metrics, logging, tracing and alerting, and define SLOs and operational health metrics that teams trust •Real-time and batch data ingestion pipelines, feature stores and data quality •Third-party connectors, APIs and platform integrations •Build model serving and inference pipelines, experiment tracking and the MLOps tooling for deployment, versioning, drift monitoring and lifecycle management Qualifications: •Hands-on experience implementing observability across APIs, cloud services and distributed systems using tools such as Prometheus, Grafana, OpenTelemetry, the ELK stack or comparable, including SLO and error-budget practice Compensation: •$185,000 - $270,000 / year •We offer performance bonuses and considerable equity Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!