Skip to content
← All jobs

Machine Learning Engineer, Voice

ElevenLabs

London, UKHybridFull-timeMid-level3+ years£95,000 - £130,000

Real-time generative audio, where latency and naturalness are both non-negotiable. You'll work across training and inference.

Requirements

  • 3+ years of ML engineering, ideally with audio or speech models
  • Experience training and serving low-latency generative models
  • Comfortable optimizing inference for real-time use cases

Responsibilities

  • Improve latency and naturalness of our real-time voice synthesis model
  • Build the evaluation pipeline for new voice model checkpoints
  • Work closely with the inference team on serving optimizations

Skills

  • Python
  • PyTorch
  • CUDA
  • Distributed systems