Wayve Logo

Wayve

SWE, Data Ingestion

Posted 2 Days Ago
Be an Early Applicant
In-Office
Sunnyvale, CA, USA
210K-250K Annually
Entry level
In-Office
Sunnyvale, CA, USA
210K-250K Annually
Entry level
Build, operate, and improve large-scale data ingestion pipelines supporting autonomous driving AI systems. Responsibilities include debugging pipeline failures, handling corrupt and inconsistent data, optimizing Spark jobs, managing orchestration and retries, improving throughput and reliability, reducing operational toil, and supporting downstream annotation, data science, training, and evaluation teams.
The summary above was generated by AI
About us   

Founded in 2017, Wayve is the leading developer of Embodied AI technology.  Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems.

Our vision is to create autonomy that propels the world forward.  Our intelligent, mapless, and hardware-agnostic AI products are designed for automakers, accelerating the transition from assisted to automated driving. 
In our fast-paced environment big problems ignite us—we embrace uncertainty, leaning into complex challenges to unlock groundbreaking solutions. We aim high and stay humble in our pursuit of excellence, constantly learning and evolving as we pave the way for a smarter, safer future.

At Wayve, your contributions matter.  We value diversity, embrace new perspectives, and foster an inclusive work environment; we back each other to deliver impact.  

Make Wayve the experience that defines your career!  

The Role

We are looking for a Data Ingestion Engineer to help build and strengthen the data foundations behind Wayve’s self-driving technology.

At Wayve, we do not hand-code cars to drive. We train them to drive from data. That makes data ingestion one of the most important parts of our learning system. The faster, more reliably and more intelligently we can process real-world driving data, the faster we can improve our models and bring embodied AI into the world.

This is a hands-on permanent role for an engineer who enjoy solving practical, high-impact problems at scale. You will help keep our ingestion pipelines running smoothly, unblock critical data flows, and contribute to the long-term evolution of the systems that support annotation, data science, model training and evaluation.

Our data platform operates at significant scale, with over 500,000 hours of driving data, equating to 100’s of PBs. As our ADAS and autonomy work grows, we need ingestion systems that are robust, efficient and cost-effective. A single bad data segment can block a pipeline, build up queues and slow down downstream teams, so this role has a direct impact on how quickly Wayve’s AI can learn.

Key Responsibilities

You will work within the Data Ingestion team to improve the reliability, efficiency and throughput of the pipelines that move real-world driving data through Wayve.

  • Debug and resolve failing or blocked ingestion pipelines.
  • Investigate issues caused by corrupt, malformed or unexpected data.
  • Design and implement more resilient pipelines so individual bad data segments do not block wider workflows.
  • Improve how we handle varied data formats from partners, suppliers and third-party sources.
  • Support orchestration across multi-step ingestion workflows, including dependencies, retries and queue management.
  • Optimise Spark jobs and data-processing pipelines for throughput, compute efficiency and reliability.
  • Reduce operational toil around failed jobs, stalled pipelines and manual interventions.
  • Work on high-volume batch-processing systems where throughput, reliability and cost all matter.
  • Help prioritise and unblock important datasets for downstream annotation, data science and model training teams.
  • Partner with engineers across Data Platform and downstream teams to deliver both immediate improvements and scalable long-term solutions.
  • Contribute to the technical direction, maintainability and operational excellence of Wayve’s ingestion platform.

About You

You are an experienced Data Engineer, Platform Engineer or Distributed Systems Engineer who enjoys working on large-scale production data systems.

You have seen how data pipelines behave in the real world: messy inputs, strange edge cases, corrupt files, stalled queues, unexpected formats and failures that only appear at scale. You are comfortable digging into those problems, finding the root cause and making systems better as a result.

You combine strong technical depth with a practical, collaborative approach. You can take ownership of complex systems, work effectively across teams and balance urgent operational needs with thoughtful, durable engineering improvements.

Essential

  • Strong production experience with Apache Spark.
  • Strong Python engineering experience.
  • Experience building, debugging or operating large-scale data-ingestion, ETL or data-processing pipelines.
  • Experience with distributed data-processing systems.
  • Ability to optimise jobs for throughput, compute efficiency and reliability.
  • Experience debugging production pipeline failures.
  • Comfort working with messy, corrupt, incomplete or inconsistent data.
  • Understanding of orchestration across multi-step pipelines and downstream dependencies.
  • Ability to work independently in a fast-moving, highly technical environment.
  • A practical, delivery-focused mindset with a focus on continuous improvement.
  • Experience working at significant data scale, ideally PB-scale or similarly high-throughput environments.

Desirable

Experience in one or more of the following areas would be a strong advantage:

  • Airflow, Flyte, Databricks Workflows or similar orchestration tooling.
  • Databricks, Delta Lake or Delta tables.
  • Scala or Java, especially in Spark-based environments.
  • Queue-based processing, retry handling and priority data workflows.
  • High-throughput batch data-processing systems.
  • Production systems with many data producers, consumers or external data sources.
  • Handling third-party, partner or supplier data with inconsistent formats and quality issues.
  • Automotive, robotics, autonomy, mapping, ML data platforms or embodied AI environments.
  • Cost optimisation for compute- and storage-heavy data platforms.
  • High-performance engineering experience from domains such as trading, where it includes relevant distributed-systems or throughput-focused work.

This is a full-time, permanent role based in our Sunnyvale, CA  office. and the reasonably estimated salary for this role ranges from $210,000 to $250,000, plus a competitive equity package. Actual compensation is based on the candidate's skills, qualifications, and experience.

At Wayve, we want the best of all worlds, so we operate a hybrid working policy that combines time together in our offices and workshops to fuel innovation, culture, relationships and learning, with time spent working from home.


At Wayve we're committed to creating a diverse, fair and respectful culture that is inclusive of everyone based on their unique skills and perspectives, and regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, veteran status, pregnancy or related condition  (including breastfeeding) or any other basis as protected by applicable law.  

For more information visit Careers at Wayve. 

To learn more about what drives us, visit Values at Wayve 

DISCLAIMER: We will not ask about marriage or pregnancy, care responsibilities or disabilities in any of our job adverts or interviews. However, we do look to capture information about care responsibilities, and disabilities among other diversity information as part of an optional DEI Monitoring form to help us identify areas of improvement in our hiring process and ensure that the process is inclusive and non-discriminatory.

Wayve is committed to creating an inclusive interview experience. If you require any accommodations or adjustments to participate fully in our interview process, please let us know.

We understand that everyone has a unique set of skills and experiences and that not everyone will meet all of the requirements listed above. If you’re passionate about self-driving cars and think you have what it takes to make a positive impact on the world, we encourage you to apply.
At Wayve we're committed to creating a diverse, fair and respectful culture that is inclusive of everyone based on their unique skills and perspectives, and regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, veteran status, pregnancy or related condition  (including breastfeeding) or any other basis as protected by applicable law.  

For more information visit Careers at Wayve. 

To learn more about what drives us, visit Values at Wayve 

For US candidates only, please visit E-Verify Notice and Participation and Right to Work

DISCLAIMER: We will not ask about marriage or pregnancy, care responsibilities or disabilities in any of our job adverts or interviews. However, we do look to capture information about care responsibilities, and disabilities among other diversity information as part of an optional DEI Monitoring form to help us identify areas of improvement in our hiring process and ensure that the process is inclusive and non-discriminatory.



Wayve Mountain View, California, USA Office

709 N Shoreline Blvd, Mountain View, California, United States, 94043

Wayve Sunnyvale, California, USA Office

605 W. California Ave, Sunnyvale, United States, 94086

Similar Jobs

3 Minutes Ago
Hybrid
Senior level
Senior level
Financial Services
Leads software engineering for secure, scalable AWS-based healthcare payment systems. Designs, develops, tests, troubleshoots, and reviews production code; builds Terraform infrastructure modules; designs resilient AWS architectures; automates CI/CD and remediation; and promotes operational stability. Leads architectural evaluations, engineering communities of practice, and responsible adoption of AI-assisted development tools, including validation for correctness, security, performance, and compliance.
Top Skills: .NetAuroraAWSBashC#Ci/CdDynamoDBEbsEc2EksEventbridgeKinesis Data StreamsLambdaLoad BalancersNaclsPowershellPythonRdsRoute 53S3Security GroupsSnsSqsSubnetsTerraformVpc
25 Minutes Ago
In-Office
San Francisco, CA, USA
150K-200K Annually
Senior level
150K-200K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Real Estate
Own paid acquisition and website conversion optimization to drive qualified pipeline and measurable growth. Manage LinkedIn, Meta, SEM, and other performance channels; oversee funnel metrics, attribution, and ROI; design A/B tests and growth experiments; improve landing pages and UX flows; analyze results using marketing and BI tools; and evolve the marketing technology stack in partnership with product, design, engineering, and GTM teams.
Top Skills: Ga4Google SearchHubspotLinkedInLookerMetaSalesforceSem
37 Minutes Ago
Hybrid
San Francisco, CA, USA
160K-210K Annually
Senior level
160K-210K Annually
Senior level
Artificial Intelligence • Hardware • Machine Learning • Robotics • Software • Utilities
Develop robotics perception systems spanning SLAM, mapping, sensor fusion, and calibration. Improve performance across varied operating conditions, including GPS-degraded environments; fuse multi-user and multi-session point-cloud maps; align LiDAR and imagery; automate sensor calibration; and improve vehicle sensor data quality, scalability, and reliability.
Top Skills: C++CamerasExtrinsic Sensor CalibrationGpsIntrinsic Sensor CalibrationLidarPoint CloudsPythonSensor FusionSlam

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account