Skip to content
Cantina logo

Machine Learning Engineer, TTS

CantinaDigital Transformation company
Europe€170K - €190KSenior
Data & AI

About the role

TL;DR

Develop and optimize large-scale speech models for Cantina's AI platform.

  • We're looking for a Research / ML Engineer to join our Speech Team to build state-of-the-art speech systems end-to-end from data specs through production inference.
  • Key Responsibilities Architect, implement, pre-train, fine-tune, and post-train/alignment for large-scale speech models.
  • Independently lead small research projects while collaborating on larger team initiatives.
  • Design, run, and analyze scientific experiments to advance our understanding of the models.
  • Develop and improve dev tooling to enhance team productivity.
  • Contribute to the entire stack, from low-level optimizations to high-level model design.
  • Requirements Exceptional research/development experience with large scale audio models (>3B models and >500k hours data).
  • Exceptional understanding and hands-on experience with transformer architectures and/or diffusion models (inc. distillation and streaming) and/or audio language modelling.
  • Strong experience with multi-node and multi-gpu distributed model training.
  • Strong software engineering skills with a proven track record of building complex systems.
  • Strong with PyTorch and performance work (profiling, CUDA/Triton/C++ as needed) and writing reliable production quality code.
View original posting →

Required skills

PyTorchC++Python

Nice-to-have skills

AWSGoogle CloudKubernetesCI/CDDocker

Domain expertise

ai

Benefits & perks

Competitive salary, Equity, Health insurance, Paid time off, Parental leave, 401(k) retirement savings plan, Lifestyle spending account, Complimentary lunch and snacks

Tech stack

PyTorchC++PythonAWSGoogle CloudKubernetesCI/CDDocker

Similar jobs

C

Machine Learning Engineer - Voice Conversion

F

Forward Deployed Engineer

F

Research Scientist - AI Safety