SWE, Data Ingestion

Wayve - Sunnyvale, CA

Hiring: SWE, Data Ingestion Company: Wayve Location: Sunnyvale, CA Job Posted Time: 2026-09-16 21:43:13 Employment Type: Full-time / Hybrid Target Skills & Keywords : Airflow, Data Pipeline, Databricks, Delta Lake, ETL, Java, Python, Scala, Segment, Spark About the job Required Skills: •This is a hands-on permanent role for an engineer who enjoy solving practical, high-impact problems at scale. You will help keep our ingestion pipelines running smoothly, unblock critical data flows, and contribute to the long-term evolution of the systems that support annotation, data science, model training and evaluation. •Work within the Data Ingestion team to improve the reliability, efficiency and throughput of the pipelines that move real-world driving data through Wayve. •Debug and resolve failing or blocked ingestion pipelines. •Investigate issues caused by corrupt, malformed or unexpected data. •Design and implement more resilient pipelines so individual bad data segments do not block wider workflows. •Improve how we handle varied data formats from partners, suppliers and third-party sources. •Support orchestration across multi-step ingestion workflows, including dependencies, retries and queue management. •Optimise Spark jobs and data-processing pipelines for throughput, compute efficiency and reliability. Qualifications: •You are an experienced Data Engineer, Platform Engineer or Distributed Systems Engineer who enjoys working on large-scale production data systems. •You have seen how data pipelines behave in the real world: messy inputs, strange edge cases, corrupt files, stalled queues, unexpected formats and failures that only appear at scale. You are comfortable digging into those problems, finding the root cause and making systems better as a result. •You combine strong technical depth with a practical, collaborative approach. You can take ownership of complex systems, work effectively across teams and balance urgent operational needs with thoughtful, durable engineering improvements. •Strong production experience with Apache Spark. •Demonstrated capacity to optimise jobs for throughput, compute efficiency and reliability. •Comfort working with messy, corrupt, incomplete or inconsistent data. •Understanding of orchestration across multi-step pipelines and downstream dependencies. •Demonstrated capacity to work independently in a fast-moving, highly technical environment. •A practical, delivery-focused mindset with a focus on continuous improvement. Compensation: •$210,000 - $250,000 / year •And the reasonably estimated salary for this role ranges from $210,000 to $250,000, plus a competitive equity package Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!