Embedded AI Engineer, On-Device Models

Deepgram · USA | Remote · Remote-friendly

Don’t apply blind. See how your CV matches Engineer first — free, in 30 seconds.

You will leave NewLuxJob. We do not receive or handle applications.

Company
Deepgram
Location
USA | Remote
Employment type
Full-time
Posted
July 7, 2026

About this job

ABOUT THE ROLE Deepgram's speech models are among the fastest and most accurate in the world, and we have deep machinery for running them on NVIDIA GPUs. Our customers need them on everything else: non-NVIDIA accelerators, embedded SoCs, mobile application processors, DSPs and NPUs, and purpose-built devices with tight memory, compute, thermal, and power budgets. When a target platform's standard kernels and runtime can't run a Deepgram model well enough, someone has to go below them. That is this role. As an Embedded AI Engineer on the Partner Platform Engineering team, you work at the lowest layer of our edge stack. You write and optimize custom kernels and operators for specific hardware, collapse models onto device-specific execution units, and do the target-side quantization and

This is a short summary.

Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.

Most-requested skills in United States

Based on 371 United States vacancies that list requirements, these are the skills employers ask for most often.

  • Project Management35% of postings
  • CAD24% of postings
  • AutoCAD21% of postings
  • MATLAB18% of postings
  • Revit12% of postings
  • Lean11% of postings
  • SolidWorks11% of postings
  • FEM/FEA7% of postings
Engineer skills

Similar jobs

See if your CV fits this job

Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.

  • Instant match score for this role
  • Tailored cover letter in one click
  • Free — no credit card
Check my CV — free
Embedded AI Engineer, On-Device Models — Deepgram | NewLuxJob