The opportunity
he Personalization team makes deciding what to play next easier and more enjoyable for every listener. From Blend to Discover Weekly, we're behind some of Spotify's most-loved features.
What you'll do
Develop and experiment with new methods for speech synthesis and speech: recognition, along with end-to-end approaches, building on the latest research and ideas.
Work towards the expansion of our speech use-cases targeting different markets and products.
Be part of a highly motivated research team dedicated to building and: creating models at scale to power the Spotify platform.
Champion best practices for research and development, sharing your knowledge: and experience with other researchers within Speak.
Collaborate with our engineering and data teams on ideas requiring new: infrastructure or new high-quality data, as well as to help improve our speech recognition and speech synthesis pipelines, and help turn proven ideas into scalable products.
What they're looking for
- You have a strong background in ML (PhD degree on top of professional experience), and
- experience in working with any of the following: transformers, GANs, diffusion models, flow matching, VAEs, audio codecs.
- You have experience in developing generative models for speech synthesis,: speech recognition, audio/music, natural language processing, or computer vision.
- You have strong experience with Python, particularly PyTorch.
- You have strong communication skills and the ability to explain technical: ideas with clarity to technical and non-technical people alike.
- You have experience in an academic or professional setting conducting high-quality research.