Research Engineer, AI Inference at ElevenLabs
Confirmed open
Deploy ElevenLabs’ AI models to production and optimize real-time serving for latency, throughput and cost. The posting names quantization, distillation, KV-cache optimization, batching and custom kernels. Remote work is limited to listed countries.
- Location
- Remote
- Advertised region
- GB / US / PL / BG
- Employment
- FULL_TIME
- Work arrangement
- remote
Role overview
Deploy ElevenLabs’ AI models to production and optimize real-time serving for latency, throughput and cost. The posting names quantization, distillation, KV-cache optimization, batching and custom kernels. Remote work is limited to listed countries.
How to apply
Review the original posting and apply through the employer’s careers page.
View source and applySource checked: 2026-10-02
Show what you can do
Build a free profile around your CV and real work. Share its link when you are ready. Applications for this role still go directly to the employer.
Build a free profile