Senior AI Inference Engineer - Model Optimization & Deployment
Zoox · Hybrid
This listing is no longer verified as available.
MeritLog keeps this source-backed description for reference. Availability is not verified, and there is no application link here.
Source: the employer's Lever job board. Open the original listing for current details. Availability is not verified for this retained page.
Job details
- Work model
- Hybrid
- Salary
- Not listed by source
- Location
- Foster City, CA; San Diego, CA; Seattle, WA
Job description
The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence. As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.
Keep exploring
Available Engineering roles
These current listings are available to explore now.
- AI Research Engineer, Post-TrainingLovable · On-site
- AI Ops Engineer (FBOS)Lovable · On-site
- Product Experience Team LeadLovable · Hybrid
- Software Engineer, Platform (Infrastructure)Lovable · On-site
- Transformation LeadLovable · On-site
- People Partner - Engineering, Product, Design (EPD)Lovable · On-site