Senior AI Inference Engineer - Model Optimization & Deployment

Zoox · Foster City

Don’t apply blind. See how your CV matches Engineer first — free, in 30 seconds.

You will leave NewLuxJob. We do not receive or handle applications.

Company
Zoox
Location
Foster City
Employment type
Full-time
Posted
April 11, 2026

About this job

The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence. As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.

This is a short summary.

Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.

Most-requested skills in United States

Based on 530 United States vacancies that list requirements, these are the skills employers ask for most often.

  • MATLAB45% of postings
  • CAD41% of postings
  • Lean15% of postings
  • SolidWorks14% of postings
  • Project Management9% of postings
  • Six Sigma9% of postings
  • FEM/FEA9% of postings
  • ISO 90017% of postings
Engineer skills

Similar jobs

See if your CV fits this job

Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.

  • Instant match score for this role
  • Tailored cover letter in one click
  • Free — no credit card
Check my CV — free
Senior AI Inference Engineer - Model Optimization & Deployment — Zoox | NewLuxJob