Embedded AI Engineer, On-Device Models at Deepgram
Confirmed open
Adapt Deepgram speech models to embedded chips, DSPs and NPUs through custom kernels and device-specific optimization.
- Location
- Remote
- Advertised region
- USA - Remote / United States
- Employment
- FullTime
- Work arrangement
- remote
- Technologies
- C++, Linux
Create a free application profile — Keep your skills and projects in a reusable profile. Private by default; you choose when to share. Apply for this job directly with the employer.
Role overview
This role writes C, C++ or Rust kernels and low-level operators for non-NVIDIA accelerators, quantizes and fuses model operations, and benchmarks latency, memory and power. It integrates vendor toolchains and delivers runtime components for embedded Linux, bare-metal or RTOS targets.
How to apply
Review the original posting and apply through the employer’s careers page.
View source and applyApply directlySource checked: 2026-10-11