Senior Software Engineer - Systems

Boson AI - Santa Clara, CA

Hiring: Senior Software Engineer - Systems Company: Boson AI Location: Santa Clara, CA Job Posted Time: 2026-09-12 07:37:44 Target Skills & Keywords : AI, AWS, Airflow, C++, CI/CD, Data Pipeline, ELT, ETL, Flink, GCP, Java, Kafka, Kinesis, Kubernetes, LLM, LangChain, LlamaIndex, Python, RAG, React, Rust, Spark About the job Experience: •3+ years building and operating backend systems at scale — you've owned services that other teams depend on in production Required Skills: •Build and operate the core platform behind Boson's model APIs and agentic products. You'll own the infrastructure that every Boson agent runs on — API serving, state management, data pipelines, context retrieval, and execution runtime — and make it fast, reliable, and easy for product teams to build on. •Own and evolve the core platform infrastructure: API serving layer, state management, policy enforcement engine, and execution runtime for agentic workflows •Design and operate high-throughput, low-latency distributed services that back our model API products — including request routing, load management, rate limiting, and multi-tenant isolation •Build and maintain downstream data pipelines (ETL/ELT) for API logs, usage analytics, and billing — ensuring data correctness, freshness, and queryability at scale •Develop production-grade internal SDKs and libraries with clean APIs, strong type safety, and clear contracts that product teams can build on confidently •Architect context and memory systems for conversational workloads — low-latency retrieval, caching, and integration with vector stores and retrieval pipelines •Instrument end-to-end observability: define SLIs/SLOs, build structured logging and tracing, and drive reliability improvements across the platform •Collaborate closely with ML and product teams to integrate model serving, voice runtime, and tooling infrastructure under tight latency and quality constraints Qualifications: •Strong distributed systems fundamentals: concurrency, fault tolerance, consistency tradeoffs, capacity planning •Applied hands-on capability in data pipeline infrastructure (Kafka/Kinesis, Spark/Flink, Airflow, or similar) for log processing, analytics, or ETL workloads •Track record of designing APIs and frameworks adopted by other engineering teams — you care about developer experience and long-term maintainability •Proficiency in at least one systems language (Go, Rust, Java, C++) or Python in a performance-sensitive context •Comfortable working across the stack: cloud infrastructure (AWS/GCP), containerized deployments (K8s), CI/CD, and production oncall •Operational familiarity with emerging agent integration protocols (MCP, A2A) or orchestration frameworks (LangChain, LlamaIndex) •Background in real-time media systems (audio/video streaming, low-latency signaling) Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!