Đăng nhậpTải ứng dụng

Việc làm › Tin đăng Mới hôm nay

Data Engineer, Application Software

Wayve · Tokyo, Japan

Toàn thời gianLàm tại chỗTiếng Anh

Về vị trí này

Từ tin đăng của nhà tuyển dụng · Wayve · đăng ngày 5 tháng 10, 2026

Before the detail, here's the challenge you'd help us solve. We build the embodied intelligence that moves real vehicles safely, and the ecosystem a billion machines will run on in the future. Very few people in AI can say this. Every role here, whatever the team, plugs into that. Here’s what this particular role covers. 🛠️ About our Engineering Teams The Machine Learning team within Application Software, part of Product & Delivery, works on critical initiatives that push the frontier of model-based autonomous driving. That covers core driving performance as well as feature-level intelligence such as personalisation, comfort and collaboration. 🧠 Your day-to-day As a Data Engineer, you’ll design and deliver scalable data pipelines that turn vast amounts of data from diverse internal and external sources into structured, reliable and model-ready datasets. Your work spans data ingestion, data quality assurance, transformation, curation, evaluation and ML support. You’ll work closely with Wayve’s Data Corpus teams, customer programmes and ML engineers, bringing up ingestion pipelines for new vehicle platforms and sensor configurations, and tackling the highest-impact bottlenecks first. 🧩 What you’ll be working on: Building and improving scalable data pipelines that support model development, evaluation and production ML workflows for autonomous driving. Ingesting, transforming and curating large-scale real-world, synthetic and partner-provided datasets into structured, reliable, model-ready formats aligned with standardised taxonomies and coordinate systems. Developing data quality checks, validation processes and monitoring so that raw vehicle data and processed datasets are high-quality, complete, consistent, traceable and fit for ML use cases. Curating and mining real-world and synthetic data to drive scenario diversity, coverage and feature-specific development. Improving pipeline performance, reliability and usability, reducing bottlenecks and increasing iteration velocity across ML development. Collaborating closely with ML engineers, Data Corpus, AI Platform and external partners so data pipelines integrate effectively with production-scale learning systems. 🙌 You should apply if: Essential Proven experience building and operating scalable data pipelines or distributed data processing systems in production environments. Strong software engineering skills in Python, with a solid foundation in maintainable, reliable, and well-tested software development practices. Proficient in SQL and PySpark, with experience using warehouse/OLAP concepts, window functions, and Spark for distributed data processing. Experience with modern data pipeline architectures, including workflow orchestration and DAG-based systems such as Airflow, Flyte, Ray, or similar. Solid understanding of robotics and automated driving data concepts, including sensor characteristics, timestamping and clock synchronisation, coordinate transformations, calibration, and ego-motion

Kỹ năng được nhắc đến

Product & Delivery

Xem điểm phù hợp của bạn cho mọi vị trí

BabZituna chấm điểm mọi công việc so với hồ sơ của bạn trên sáu tiêu chí thực tế và cho bạn thấy TẠI SAO lại có điểm đó, đã được kiểm định về tính công bằng (đọc báo cáo kiểm định thiên lệch công khai).

Tải ứng dụng → ✓ Miễn phí 100% cho người tìm việc
Một công việc được chấm điểm thế nào Ví dụ
Kỹ năng96Kinh nghiệm90Địa điểm84Hình thức làm việc74Loại công việc61Mức lươngKhông có dữ liệu

Số liệu minh họa, không phải ứng viên thật. Mỗi tiêu chí được chấm trên thang 100 từ chính hồ sơ của bạn, và tiêu chí nào chúng tôi không đo được sẽ ghi rõ thay vì phỏng đoán.