Machine Learning Engineer, TTS
CantinaDigital Transformation company
Europe€170K - €190KSenior
Data & AI
About the role
TL;DR
Develop and optimize large-scale speech models for Cantina's AI platform.
- •We're looking for a Research / ML Engineer to join our Speech Team to build state-of-the-art speech systems end-to-end from data specs through production inference.
- •Key Responsibilities Architect, implement, pre-train, fine-tune, and post-train/alignment for large-scale speech models.
- •Independently lead small research projects while collaborating on larger team initiatives.
- •Design, run, and analyze scientific experiments to advance our understanding of the models.
- •Develop and improve dev tooling to enhance team productivity.
- •Contribute to the entire stack, from low-level optimizations to high-level model design.
- •Requirements Exceptional research/development experience with large scale audio models (>3B models and >500k hours data).
- •Exceptional understanding and hands-on experience with transformer architectures and/or diffusion models (inc. distillation and streaming) and/or audio language modelling.
- •Strong experience with multi-node and multi-gpu distributed model training.
- •Strong software engineering skills with a proven track record of building complex systems.
- •Strong with PyTorch and performance work (profiling, CUDA/Triton/C++ as needed) and writing reliable production quality code.
Required skills
PyTorchC++Python
Nice-to-have skills
AWSGoogle CloudKubernetesCI/CDDocker
Domain expertise
ai
Benefits & perks
Competitive salary, Equity, Health insurance, Paid time off, Parental leave, 401(k) retirement savings plan, Lifestyle spending account, Complimentary lunch and snacks
Tech stack
PyTorchC++PythonAWSGoogle CloudKubernetesCI/CDDocker