Senior AI Inference Engineer - Model Optimization & Deployment

Zoox · Foster City, CA · $225k - $305k
full-time senior Posted 4 months ago

Before you apply

Build my evidence-backed draft — free Apply on company site →

Paste your relevant resume section or 2–4 true bullets. See supported requirements and honest gaps. No account and no application sent.

Get weekly job alerts like this →

About this role

The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence. As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.

Similar Jobs

Related searches:

On-site Jobs Senior Jobs On-site Senior Jobs Senior AI InfrastructureSenior NLP & Language AISenior Generative AISenior Machine Learning AI Jobs in Foster City AI Infrastructure in Foster CityNLP & Language AI in Foster CityGenerative AI in Foster CityMachine Learning in Foster City generative-aigpullminference

Get jobs like this delivered weekly

Free AI jobs newsletter. No spam.