Senior AI Inference Engineer - Model Optimization & Deployment
Zoox
| Company | Zoox |
| Category | Engineering |
| Location | Foster City |
| Remote | Hybrid |
| Employment | Full-time |
| Level | Senior |
| Salary | USD 242k–290k |
| Posted | 11 Apr 2026 |
| Last verified | 31 Jul 2026 |
| Source | Employer career page (lever) |
Description
The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence.
As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.
1,064,721 openings. Erioun finds yours.Scored against your own profile, every hour.Try the radar →