About the Role
Join our Artist-First AI Music Lab as a Research Scientist to pioneer and advance state-of-the-art generative technologies for music, creating breakthrough experiences for artists and fans. We invent new listening experiences that center and celebrate artists and creatives, guided by principles of artist partnerships, choice in participation, fair compensation, and deepening artist-fan connections.
Responsibilities
- Conduct groundbreaking research in generative audio using diffusion or flow matching models, focusing on vocal synthesis, post-training alignment techniques, or iterative music generation and audio editing.
- Run large-scale experiments utilizing Spotify's infrastructure and reaching over 700 million monthly active users.
- Create practical applications that leverage generative technologies to push the boundaries of listening experiences.
- Collaborate within a cross-functional team of scientists, engineers, product managers, designers, user researchers, and analysts.
- Impact Spotify's products, tools, and services with projects influencing the entire organization.
- Engage with the research community through publications, talks, and conference attendance.
Requirements
- Ph.D. in Computer Science, Mathematics, Engineering, or a related field.
- Experience in generative modeling, machine learning, music information retrieval, speech processing, audio processing, signal processing, probabilistic modeling, computer vision, or related areas.
- Deep expertise in vocal/speech synthesis, post-training alignment techniques (e.g., PPO, GRPO, DPO), or audio-to-audio generation and text-guided music editing.
- Publications at leading conferences such as ICASSP, ISMIR, INTERSPEECH, ICLR, AAAI, IJCAI, NeurIPS, ICML, CVPR, ECCV, ICCV, or related venues.
- Strong coding skills in Python, PyTorch, and NumPy.
- Creative problem-solving abilities with a passion for building outstanding products.
- Enthusiasm for translating research ideas into scalable products.
- Ability to explain complex topics simply and build strong relationships with colleagues and stakeholders.
- Previous industry experience is helpful.
Skills
- Generative modeling
- Machine learning
- Music information retrieval
- Speech processing
- Audio processing
- Signal processing
- Probabilistic modeling
- Computer vision
- Python
- PyTorch
- NumPy
- Diffusion models
- Flow matching models
- Vocal synthesis
- Speech synthesis
- Post-training techniques
- Preference alignment methods
- Reinforcement learning
- Iterative music generation
- Audio editing
Location
- North Americas region
Work Type
- Remote
- Hybrid
Experience Level
- All levels of seniority
Education Level
- Ph.D.
Salary/Compensations
- $133,194 - $190,278 plus equity
Benefits
- Health insurance
- Six month paid parental leave
- 401(k) retirement plan
- Monthly meal allowance
- 23 paid days off
- 13 paid flexible holidays
About the Company
- Spotify is an equal opportunity employer.
- Our platform is for everyone, and so is our workplace.
- We are passionate about inclusivity and making sure our entire recruitment process is accessible to everyone.
- We have ways to request reasonable accommodations during the interview process and help assist in what you need.
Equal Opportunity
- You are welcome at Spotify for who you are, no matter where you come from, what you look like, or what’s playing in your headphones.
- The more voices we have represented and amplified in our business, the more we will all thrive, contribute, and be forward-thinking!
- So bring us your personal experience, your perspectives, and your background.
- It’s in our differences that we will find the power to keep revolutionizing the way the world listens.
- If you need accommodations at any stage of the application or interview process, please let us know - we’re here to support you in any way we can.
