Software Engineer, Data Flywheel Platform
full-time
mid
Posted 20 hours ago
Before you apply
Build my evidence-backed draft — free Apply on company site →Paste your relevant resume section or 2–4 true bullets. See supported requirements and honest gaps. No account and no application sent.
About this role
About us
Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems.
Our vision is to create autonomy that propels the world forward. Our intelligent, mapless, and hardware-agnostic AI products are designed for automakers, accelerating the transition from assisted to automated driving. In our fast-paced environment big problems ignite us—we embrace uncertainty, leaning into complex challenges to unlock groundbreaking solutions. We aim high and stay humble in our pursuit of excellence, constantly learning and evolving as we pave the way for a smarter, safer future.
At Wayve, your contributions matter. We value diversity, embrace new perspectives, and foster an inclusive work environment; we back each other to deliver impact.
Make Wayve the experience that defines your career!
About Wayve and the team
Wayve is building embodied AI for the physical world , starting with autonomous driving. Instead of the hand-engineered, modular stacks that defined the first era of self-driving, we pioneered AV2.0: a single, end-to-end neural network that learns to drive from raw sensor data and generalizes to new cities, vehicles, and conditions. Our foundation models, the GAIA family of generative world models and the LINGO family of vision-language-action models, allow vehicles to perceive, reason, and act in the open world. We have driven zero-shot across hundreds of cities on three continents, and we are now scaling from proving the science to deploying it with leading automakers and mobility partners, including Nissan, Stellantis, and Uber.
This role sits in the AI Platform organization, on the data flywheel that powers every model we ship. Applied Scientists and ML Engineers on the team push the frontier on data curation, enrichment, foundation-model evaluation, and the models themselves. This role builds the platform underneath all of it : the pipelines, infrastructure, and systems that turn world-scale fleet data into high-signal training data, evaluate and train foundation models, and enable every team to run these workflows themselves. As deployment scales, the leverage is enormous: the better the platform, the faster the whole flywheel turns.
The role
We are hiring a senior Software Engineer to build the platform that powers Wayve’s data flywheel and foundation-model stack. This is the engineering counterpart to our Applied Scientist and ML Engineer roles : you build the systems they, and the wider company, depend on. It is high-leverage, high-visibility work with a clear path to deep system ownership.
Build the systems that allow teams to turn world-scale driving data into high-signal training data, and evaluate and train foundation models on it.
Replace ad-hoc scripts and manual handoffs with self-serve, observable products used across Science, Autonomy, and Evaluation.
Every model Wayve ships runs on this platform: your work compounds across the entire fleet and roadmap.
Work shoulder to shoulder with a world-class science and engineering team, with real deployment at global OEM scale (Nissan, Stellantis, Uber).
TC3 / TC4 ownership of platform and infrastructure, with room to set technical direction as the platform matures.
What you will do
Build and scale the data curation and enrichment pipelines that turn world-scale fleet data into high-signal training data: mining and active-learning loops, running model-based enrichments over billions of rows, and ensuring data quality at scale.
Build the evaluation infrastructure behind foundation-model progress: harnesses for offline and closed-loop evaluation, metric and benchmark pipelines, and world-model-based evaluation.
Build and optimize training and serving infrastructure for large pretrained models: distributed training, batched inference, and large-scale model backfills.
Build the data-platform backbone: distributed data processing (Ray Data, Daft, Spark / Databricks), embedding and vector search (turbopuffer, Milvus), lakehouse formats (Lance, Iceberg), dataset versioning, and the enrichment and annotation catalog.
Make it self-serve and reliable: turn one-off processes into products that other teams operate themselves, and own testing, observability, and on-call for what you ship.
Partner closely with Applied Scientists and ML Engineers to take research from prototype to production at scale.
What we are looking for
Strong production software engineering , especially production Python (services, APIs, large-scale data processing), and comfort owning and extending large codebases.
Large-scale data and distributed-systems experience: batch and streaming pipelines, workflow orchestration (Flyte, Airflow, Dagster, or similar), and distributed processing (Spark / PySpark, Ray, Databricks, or e
Similar Jobs
Related searches:
On-site Jobs
Mid-Level Jobs
On-site Mid-Level Jobs
Mid-Level AI ResearchMid-Level Generative AIMid-Level Machine LearningMid-Level Computer VisionMid-Level AI InfrastructureMid-Level Robotics & AutonomyMid-Level Backend & Systems
AI Jobs in Sunnyvale
AI Research in SunnyvaleGenerative AI in SunnyvaleMachine Learning in SunnyvaleComputer Vision in SunnyvaleAI Infrastructure in SunnyvaleRobotics & Autonomy in SunnyvaleBackend & Systems in Sunnyvale
roboticsautonomous-vehiclesmlopsdistributed-systemsgenerative-aideep-learningevaluation
Get jobs like this delivered weekly
Free AI jobs newsletter. No spam.