Spotify

Senior Applied Research Scientist - Personalization

Spotify • $169K — $241K *
US-AnywhereRemote in New York, NY
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • PhD in ML or related field with professional experience.
  • Proficiency in developing ML techniques like transformers and GANs.
  • Experience with generative models in speech synthesis and recognition.
  • Strong programming skills in Python, especially with PyTorch.
  • Excellent communication skills for diverse audiences.

Responsibilities

  • Develop and test new methods for speech synthesis and recognition.
  • Expand speech applications for various markets and products.
  • Collaborate with a motivated research team to scale models for Spotify.
  • Promote best practices in research and share knowledge with colleagues.
  • Work with engineering and data teams to enhance infrastructure and processes.

Benefits

  • Health insurance coverage.
  • Six months paid parental leave.
  • 401(k) retirement plan to save for the future.
  • Monthly meal allowance to support your daily needs.
  • 23 paid days off for vacation and personal time.
  • 13 paid flexible holidays to celebrate special occasions.
Full Job Description
Within Personalization, the Speak Team owns the development of Spotify's state-of-the-art speech models, contributing to speech recognition, speech synthesis, and speech-to-speech models. We craft voice models that match human-level emotional expressiveness, so we can deeply engage our listeners and support creators at scale. Our groundbreaking work on speech synthesis relies on state-of-the-art deep learning methods and evaluation techniques, highly efficient data processing and model serving, and capturing audio of outstanding quality from our voice talent pool.

We're looking for a senior applied research scientist with experience in developing novel ML techniques and architectures and with a strong interest in working across a full production pipeline to produce state-of-the-art generative conversational speech-to-speech models. You'll collaborate with our engineering teams to help develop our production pipelines, explore new ideas and methods to improve quality, understanding and realism, as well as push the frontiers of what is possible with our speech technology.

What You'll Do

  • Develop and experiment with new methods for speech synthesis and speech recognition, along with end-to-end approaches, building on the latest research and ideas.
  • Work towards the expansion of our speech use-cases targeting different markets and products.
  • Be part of a highly motivated research team dedicated to building and creating models at scale to power the Spotify platform.
  • Champion best practices for research and development, sharing your knowledge and experience with other researchers within Speak.
  • Collaborate with our engineering and data teams on ideas requiring new infrastructure or new high-quality data, as well as to help improve our speech recognition and speech synthesis pipelines, and help turn proven ideas into scalable products.


Who You Are

  • You have a strong background in ML (PhD degree on top of professional experience), and
  • experience in working with any of the following: transformers, GANs, diffusion models, flow matching, VAEs, audio codecs.
  • You have experience in developing generative models for speech synthesis, speech recognition, audio/music, natural language processing, or computer vision.
  • You have strong experience with Python, particularly PyTorch.
  • You have strong communication skills and the ability to explain technical ideas with clarity to technical and non-technical people alike.
  • You have experience in an academic or professional setting conducting high-quality research.


Where You'll Be

  • This role is based in New York City.
  • We offer you the flexibility to work where you work best! There will be some in person meetings, but still allows for flexibility to work from home


The United States base range for this position is $169,157 - $241,653 plus equity. The benefits available for this position include health insurance, six month paid parental leave, 401(k) retirement plan, a monthly meal allowance, 23 paid days off, 13 paid flexible holidays. These ranges may be modified in the future.

About Spotify

Spotify is a Swedish audio streaming and media services provider, launched in October 2008. The platform is owned by Spotify AB, a publicly traded company on the New York Stock Exchange since April 2018. Spotify's primary business is providing an audio streaming platform, with the company claiming that it had 345 million active monthly users and 155 million paying subscribers as of December 2020. Unlike physical or download sales, which pay artists a fixed price per song or album sold, Spotify pays royalties based on the number of artist streams as a proportion of total songs streamed on the service. Spotify distributes approximately 70% of its total revenue to rights holders, who then pay artists based on their individual agreements. Spotify has faced criticism from artists and producers including Taylor Swift and Thom Yorke, who have argued that it does not fairly compensate musicians. Spotify has also faced criticism from artists and producers including Taylor Swift and Thom Yorke, who have argued that it does not fairly compensate musicians.
Learn more about Spotify
Size
3,456 employees
Industry
Founded
2006

Similar Jobs

More Jobs at Spotify

More Information Technology Jobs

Find similar Senior Applied Research Scientist - Personalization jobs: