Logo for Spotify

Senior Applied Research Scientist - Personalization

Role overview

Qualifications

  • PhD degree with professional experience in ML
  • Experience with transformers, GANs, diffusion models, VAEs, audio codecs
  • Experience developing generative models for speech synthesis and recognition
  • Strong experience with Python, particularly PyTorch

Responsibilities

  • Develop and experiment with new methods for speech synthesis and recognition
  • Expand speech use-cases targeting different markets and products
  • Build and create models at scale to power the Spotify platform
  • Collaborate to improve speech recognition and synthesis pipelines and turn proven ideas into scalable products

Key facts

  • Remote from: New York (USA)
  • Full time
  • Senior (5-10 years)
  • Research Scientist
  • English

Hard skills

Other skills

  • Communication

About the company

Spotify logo

Spotify

Streaming Services (SVOD/AVOD)

Our mission is to unlock the potential of human creativity—by giving a million creative artists the opportunity to live off their art and billions of fans the opportunity to enjoy and be inspired by it. Spotify transformed music listening forever when it launched in Sweden in 2008. Discover, manage and share over 70m tracks for free, or upgrade to Spotify Premium to access exclusive features including offline mode, improved sound quality, and an ad-free music listening experience. Today, Spotify is the most popular global audio streaming service with 365m users, including 165m subscribers across 178 markets. We are the largest driver of revenue to the music business today.

Company details

Company typeXLarge
IndustryStreaming Services (SVOD/AVOD)
Company size5001 - 10000

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

he Personalization team makes deciding what to play next easier and more enjoyable for every listener. From Blend to Discover Weekly, we're behind some of Spotify's most-loved features. We built them by understanding the world of music and podcasts better than anyone else. Join us and you'll keep millions of users listening by making great recommendations to each and every one of them.
Within Personalization, the Speak Team owns the development of Spotify's state-of-the-art speech models, contributing to speech recognition, speech synthesis, and speech-to-speech models. We craft voice models that match human-level emotional expressiveness, so we can deeply engage our listeners and support creators at scale. Our groundbreaking work on speech synthesis relies on state-of-the-art deep learning methods and evaluation techniques, highly efficient data processing and model serving, and capturing audio of outstanding quality from our voice talent pool.
We're looking for a senior applied research scientist with experience in developing novel ML techniques and architectures and with a strong interest in working across a full production pipeline to produce state-of-the-art generative conversational speech-to-speech models. You'll collaborate with our engineering teams to help develop our production pipelines, explore new ideas and methods to improve quality, understanding and realism, as well as push the frontiers of what is possible with our speech technology.
What You'll Do
  • Develop and experiment with new methods for speech synthesis and speech recognition, along with end-to-end approaches, building on the latest research and ideas.
  • Work towards the expansion of our speech use-cases targeting different markets and products.
  • Be part of a highly motivated research team dedicated to building and creating models at scale to power the Spotify platform.
  • Champion best practices for research and development, sharing your knowledge and experience with other researchers within Speak.
  • Collaborate with our engineering and data teams on ideas requiring new infrastructure or new high-quality data, as well as to help improve our speech recognition and speech synthesis pipelines, and help turn proven ideas into scalable products.

  • Who You Are
  • You have a strong background in ML (PhD degree on top of professional experience), and
  • experience in working with any of the following: transformers, GANs, diffusion models, flow matching, VAEs, audio codecs.
  • You have experience in developing generative models for speech synthesis, speech recognition, audio/music, natural language processing, or computer vision.
  • You have strong experience with Python, particularly PyTorch.
  • You have strong communication skills and the ability to explain technical ideas with clarity to technical and non-technical people alike.
  • You have experience in an academic or professional setting conducting high-quality research.

  • Where You'll Be
  • This role is based in New York City.
  • We offer you the flexibility to work where you work best! There will be some in person meetings, but still allows for flexibility to work from home
  • The United States base range for this position is $169,157 - $241,653 plus equity. The benefits available for this position include health insurance, six month paid parental leave, 401(k) retirement plan, a monthly meal allowance, 23 paid days off, 13 paid flexible holidays. These ranges may be modified in the future.
    Spotify is an equal opportunity employer. You are welcome at Spotify for who you are, no matter where you come from, what you look like, or what’s playing in your headphones. Our platform is for everyone, and so is our workplace. The more voices we have represented and amplified in our business, the more we will all thrive, contribute, and be forward-thinking! So bring us your personal experience, your perspectives, and your background. It’s in our differences that we will find the power to keep revolutionizing the way the world listens.   At Spotify, we are passionate about inclusivity and making sure our entire recruitment process is accessible to everyone. We have ways to request reasonable accommodations during the interview process and help assist in what you need. If you need accommodations at any stage of the application or interview process, please let us know - we’re here to support you in any way we can.  

    Apply once. Then go straight to the hiring manager.

    After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

    MR

    Marcus Rivera

    Chief Revenue Officer

    m.rivera@company.com
    linkedin.com/in/marcusrivera
    Unlocked after you apply
    ·

    Research Scientist Related jobs

    Other jobs at Spotify

    Premium

    Reach out to the hiring manager directly.

    Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

    • Full match report with fit score and gaps
    • Career diagnostics on how recruiters read you
    • Curated company matches and warm intros
    • 48h early access to new roles

    Cancel anytime.