← All jobs

Research Engineer, AI Inference at ElevenLabs

Confirmed open

Deploy ElevenLabs’ AI models to production and optimize real-time serving for latency, throughput and cost. The posting names quantization, distillation, KV-cache optimization, batching and custom kernels. Remote work is limited to listed countries.

ElevenLabs initialsElevenLabs
Location
Remote
Advertised region
GB / US / PL / BG
Employment
FULL_TIME
Work arrangement
remote

Role overview

Deploy ElevenLabs’ AI models to production and optimize real-time serving for latency, throughput and cost. The posting names quantization, distillation, KV-cache optimization, batching and custom kernels. Remote work is limited to listed countries.

How to apply

Review the original posting and apply through the employer’s careers page.

View source and apply

Source checked: 2026-10-02

Show what you can do

Build a free profile around your CV and real work. Share its link when you are ready. Applications for this role still go directly to the employer.

Build a free profile

Related jobs