- Remoto
- Híbrido
- Presencial
- Tiempo completo
US$125,000 - US$165,000 a year
What you'll do We are looking for an MLOps Engineer to build and scale the inference infrastructure for our generative audio models, including Text-to-Speech (TTS), voice conversion, and Automatic Speech Recognition (ASR). You will be responsible for designing and deploying high-performance systems that ensure low-latency, ...
- Remoto
- Híbrido
- Presencial
- Tiempo completo
US$200,000 - US$220,000 a year
... speech systems end-to-end-from data specs through production inference. You'll drive the model ↔ data ↔ eval flywheel for VC and adjacent tasks (controllable TTS, voice design and more), partnering closely with research, data, and infra to ship fast, reliable, and cost-aware models. In this role, you will work at the intersection ...
- Remoto
- Híbrido
- Presencial
- Tiempo completo
US$200,000 - US$220,000 a year
... includes voice cloning and multi-speaker conditioning inside joint AV models, cinematic dialogue with music and sound design, and adjacent speech tasks (controllable TTS, voice conversion) that feed the same stack. You'll drive the model ↔ data ↔ eval flywheel, partnering closely with research, video, data, and infra to ship fast, ...