Senior AI Inference Engineer - Model Optimization & Deployment
full-time
senior
Posted 4 months ago
Before you apply
Build my evidence-backed draft — free Apply on company site →Paste your relevant resume section or 2–4 true bullets. See supported requirements and honest gaps. No account and no application sent.
About this role
The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence.
As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.
Similar Jobs
Related searches:
On-site Jobs
Senior Jobs
On-site Senior Jobs
Senior AI InfrastructureSenior NLP & Language AISenior Generative AISenior Machine Learning
AI Jobs in Foster City
AI Infrastructure in Foster CityNLP & Language AI in Foster CityGenerative AI in Foster CityMachine Learning in Foster City
generative-aigpullminference
Get jobs like this delivered weekly
Free AI jobs newsletter. No spam.