Wayve
Posted 3w ago

Data Engineer

Wayve
Leonberg, Baden-Württemberg, Germany
HybridFull Time
Responsibilities
  • building pipelines
  • ingesting data
  • curating data
Requirements
  • Proven experience building scalable production data pipelines
  • Strong Python, SQL and PySpark skills
  • Experience with DAG orchestration (Airflow, Flyte, Ray)
  • Understanding of autonomous driving data and ML workflows
Technical tools mentioned
PythonSQLPySparkAirflowFlyteRay

Job description

About us   

Founded in 2017, Wayve is the leading developer of Embodied AI technology.  Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems.

Our vision is to create autonomy that propels the world forward.  Our intelligent, mapless, and hardware-agnostic AI products are designed for automakers, accelerating the transition from assisted to automated driving. 

In our fast-paced environment big problems ignite us—we embrace uncertainty, leaning into complex challenges to unlock groundbreaking solutions. We aim high and stay humble in our pursuit of excellence, constantly learning and evolving as we pave the way for a smarter, safer future.

At Wayve, your contributions matter.  We value diversity, embrace new perspectives, and foster an inclusive work environment; we back each other to deliver impact.  

Make Wayve the experience that defines your career!  

The Role 

As a Data Engineer within the Machine Learning team in Application Software, you’ll contribute to critical initiatives that push the frontier of model-based autonomous driving—both in terms of core driving performance and feature-level intelligence such as personalization, comfort, and collaboration.

You’ll design and deliver scalable data pipelines that transform vast amounts of data from diverse internal and external sources into structured, reliable, and model-ready datasets. Your work will span data ingestion, data quality assurance, transformation, curation, evaluation and ML support. You’ll collaborate deeply with Wayve’s Data Corpus teams and ML engineers to build systems that are performant, adaptable, and ready for production.

Key Responsibilities:

  • Build and improve scalable data pipelines that support model development, evaluation, and production ML workflows for autonomous driving.
  • Ingest, transform, and curate large-scale real-world, synthetic, and partner-provided datasets into structured, reliable, and model-ready formats aligned with standardised taxonomies and coordinate systems.
  • Develop data quality checks, validation processes, and monitoring to ensure both raw data from our vehicle platforms and processed datasets are high-quality, complete, consistent, traceable, and fit for ML use cases.
  • Curate and mine real-world and synthetic data to drive scenario diversity, coverage, and feature-specific development.
  • Improve pipeline performance, reliability, and usability, helping reduce bottlenecks and increase iteration velocity across ML development.
  • Collaborate closely with Machine Learning engineers, Data Corpus, AI Platform, and external partners to ensure data pipelines integrate effectively with production-scale learning systems.

About You 

In order to set you up for success as a Data Engineer at Wayve, we’re looking for the following skills and experience.  

Essential 

  • Proven experience building and operating scalable data pipelines or distributed data processing systems in production environments.
  • Strong software engineering skills in Python, with a solid foundation in maintainable, reliable, and well-tested software development practices.
  • Proficient in SQL and PySpark, with experience using warehouse/OLAP concepts, window functions, and Spark for distributed data processing.
  • Experience with modern data pipeline architectures, including workflow orchestration and DAG-based systems such as Airflow, Flyte, Ray, or similar.
  • Solid understanding of robotics and automated driving data concepts, including sensor characteristics, timestamping and clock synchronisation, coordinate transformations, calibration, and ego-motion signals such as GNSS/IMU and vehicle odometry.
  • Understanding of machine learning development workflows, including training data generation, evaluation datasets, scenario mining, and model iteration.
  • Excellent communication and collaborative skills, capable of working effectively with interdisciplinary teams.

Desirable 

  • Prior work in calibration, perception, imitation learning, or trajectory prediction.
  • Experience with third-party dataset ingestion and transformation.
  • Familiarity with automated driving data and scenario taxonomies, including ODD definitions, manoeuvre and behavior labels, scene classification, event tagging, semantic understanding.

This is a full-time role based in our office in Leonberg. At Wayve we want the best of all worlds so we operate a hybrid working policy that combines time together in our offices and workshops to fuel innovation, culture, relationships and learning, and time spent working from home. We operate core working hours so you can determine the schedule that works best for you and your team.  

#LI-KM1

Wayve is committed to creating an inclusive interview experience. If you require any accommodations or adjustments to participate fully in our interview process, please let us know.

We understand that everyone has a unique set of skills and experiences and that not everyone will meet all of the requirements listed above. If you’re passionate about self-driving cars and think you have what it takes to make a positive impact on the world, we encourage you to apply.

At Wayve we're committed to creating a diverse, fair and respectful culture that is inclusive of everyone based on their unique skills and perspectives, and regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, veteran status, pregnancy or related condition  (including breastfeeding) or any other basis as protected by applicable law.  

For more information visit Careers at Wayve. 

To learn more about what drives us, visit Values at Wayve 

For US candidates only, please visit E-Verify Notice and Participation and Right to Work


DISCLAIMER: We will not ask about marriage or pregnancy, care responsibilities or disabilities in any of our job adverts or interviews. However, we do look to capture information about care responsibilities, and disabilities among other diversity information as part of an optional DEI Monitoring form to help us identify areas of improvement in our hiring process and ensure that the process is inclusive and non-discriminatory.

 

 

About Wayve

Develops AI software for autonomous vehicle navigation.

Year founded
2017
Employees
1000
Organization type
Private
Latest investment
Raised $1.50B Series D (2026) — led by Eclipse, Balderton, SoftBank Vision Fund 2
Headquarters
GB

Similar jobs

Data Engineer roles near Leonberg, Baden-Württemberg
12h
Save
Mark Applied
Hide
Data Engineer - Axion
Sunnyvale or Washington or San Diego or Fort Walton Beach or Ann Arbor or London or Stuttgart or Munich or Stockholm or Bangalore or Seoul or Tokyo
$200k-$275k/yr OnsiteFull Time
Applied Intuition
Applied Intuition: Developing software and simulation infrastructure for autonomous vehicles.
5+ YOERequires 5+ years of relevant experience, modern ML infrastructure, large-scale GPU jobs, data software, microservices or databases, U.S. citizenship, and eligibility for security clearance.
React, TypeScript, Python, Golang, Docker, Kubernetes, Opensearch, Postgres, LLMs, VLMs, GPU
1d
Save
Mark Applied
Hide
Senior Data Engineer (Databricks) (m/w/d)
Stuttgart, Baden-Württemberg, Germany
€62k-€72k/yr HybridFull Time
CBTW
CBTW: Global technology services firm designing and building digital solutions.
Degree in a relevant quantitative or computing field; several years developing Databricks, Spark, SQL, and cloud data solutions; Python, Scala, Terraform, CI/CD; fluent German and English; DACH travel readiness.
Databricks, Apache Spark, Kafka, Azure Data Factory, Unity Catalog, AI/BI Genie, Python, Scala, SQL, Terraform, CI/CD, MLOps, DBT, Infrastructure-as-Code (IaC)
1w
Save
Mark Applied
Hide
Data Engineer
Luxembourg or Karlsruhe
RemoteFull Time
R3 Robotics
R3 Robotics: Developing AI-powered robotic systems for automated industrial dismantling.
2+ YOE2+ years experience in data engineering or related software roles; experienced level.
1w
Save
Mark Applied
Hide
Data Engineer (m/w/d) – Simulation, Manufacturing & KI
Leinfelden-Echterdingen, Baden-Württemberg, Germany
HybridFull Time
iFAKT
iFAKT: Provider of industrial software for process simulation and optimization.
Bachelor degree in a quantitative field, strong Python and ML skills, experience with data pipelines, SQL, Git and Docker, fluent German and English, analytical and teamwork skills.
Python, Git, GitLab, Docker, CI/CD, C#, JavaScript, SQL
2w
Save
Mark Applied
Hide
Data Engineer - AFRICOM
Stuttgart, Baden-Württemberg, Germany
OnsiteFull Time
Agile Defense
Agile Defense: Provides IT modernization and cybersecurity for federal government agencies.
4+ YOETop Secret/SCI clearance required. 4+ years experience in data science/engineering, proficiency with Python, SQL, Spark/Databricks, ML model development and deployment, and working in secure cloud/containerized environments.
Python, SQL, Spark, Databricks, PySpark, scikit-learn, TensorFlow, XGBoost, MLflow, AWS, Azure, Palantir Foundry, Tableau, Plotly, Matplotlib
2w
Save
Mark Applied
Hide
Data Engineer - Bad Friedrichshall (m/w/d)
Bad Friedrichshall, Baden-Württemberg, Germany
OnsiteFull Time
Schwarz Group
Schwarz Group: International retail group operating Lidl, Kaufland, and cloud services.
4+ YOE4+ years data engineering experience; strong Spark/Databricks, Python; Apache Airflow, cloud (GCP) experience; knowledge of Iceberg/Delta Lake, Unity Catalog, Kubernetes and data security/cataloging.
spark, Databricks, Python, Scala, Apache Airflow, GCP, Apache Iceberg, Delta Lake, Unity Catalog, Kubernetes, Google Pub/Sub, CI/CD, DevOps
3w
Save
Mark Applied
Hide
Data Engineer (m/w/d) | DSIANA
Karlsruhe or Münster
€70k-€100k/yr HybridFull Time, Part Time
Atruvia
Atruvia: Provides core banking software and IT infrastructure services.
Completed university degree, strong Python and SQL skills, experience with data modeling, ETL and Big Data tools (Spark,Kafka) preferred, Openshift willingness, German C1 and English B2.
Python, SQL, Apache Kafka, Spark, Delta, S3, Openshift, Java, Java - Hibernate, Spring Framework, Angular (Web-Framework), JavaScript, TypeScript, HTML, npm (Software), Apache Maven, Git, GitLab, Jira, Integrated Development Environment (IDE), REST API, API-Design, DevSecOps
4w
Save
Mark Applied
Hide
Data Engineer
Stuttgart, Baden-Württemberg, Germany
$99k-$225k/yr OnsiteFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Consulting and technology services for government and commercial clients
3+ YOE3+ years application development and 3+ years with Python or TypeScript; active Secret clearance; Bachelor's in computer science; experience with ML, model lifecycle, and strong analytical skills.
Python, TypeScript, Maven Smart System (MSS), War Data Platform (WDP)