← All jobs
Machine Learning Engineer, Voice
ElevenLabs
London, UKHybridFull-timeMid-level3+ years£95,000 - £130,000
Real-time generative audio, where latency and naturalness are both non-negotiable. You'll work across training and inference.
Requirements
- 3+ years of ML engineering, ideally with audio or speech models
- Experience training and serving low-latency generative models
- Comfortable optimizing inference for real-time use cases
Responsibilities
- Improve latency and naturalness of our real-time voice synthesis model
- Build the evaluation pipeline for new voice model checkpoints
- Work closely with the inference team on serving optimizations
Skills
- Python
- PyTorch
- CUDA
- Distributed systems