ML Engineer, TTS & Voice AI
Summary
Build and deploy text-to-speech and voice AI models end-to-end, from data pipelines to production inference, focusing on voice cloning and controllable TTS.
Cantina is seeking a Research / ML Engineer for our Speech Team to build state-of-the-art speech systems end-to-end, from data specs to production inference. You’ll drive the model data eval flywheel for TTS and related tasks, partnering with research, data, and infra to ship reliable, cost-aware models.
You’ll lead small research projects, design experiments, and develop tooling, contributing to safety and responsible AI while pushing the boundaries of voice cloning and controllable TTS.
