Research Scientist - Generative Audio

New
S
SpotifyMusic Technology
Within the EMEA region as long as we have a work location, Central European and GMT; Core working hours are CET 3pm-6pm / EST 9am-12pm.Full-Time
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Required Skills
PythonMachine LearningNumpyPyTorch

Requirements

  • Ph.D. in Computer Science, Mathematics, Engineering, or a related field.
  • Experience in generative modeling, machine learning, music information retrieval, speech, audio, or signal processing.
  • Deep expertise in vocal/speech synthesis, post-training alignment (e.g., PPO, GRPO, DPO), or audio-to-audio generation and text-guided music editing.
  • Strong programming proficiency in Python, PyTorch, and NumPy.
  • Proven track record of research publications at leading venues such as NeurIPS, ICML, ICASSP, ISMIR, or INTERSPEECH.
  • Ability to turn research ideas into scalable, real-world product applications.
  • Excellent communication skills for explaining complex technical topics.
  • Demonstrated ability to build relationships with colleagues and stakeholders.

Responsibilities

  • Conduct groundbreaking research in generative audio using diffusion or flow matching models.
  • Research in vocal and speech synthesis, post-training alignment techniques, or iterative music generation and editing.
  • Run large-scale experiments utilizing extensive infrastructure and user data.
  • Create practical applications that push the boundaries of music listening experiences.
  • Collaborate within a cross-functional team of scientists, engineers, product managers, and researchers.
  • Publish research findings, deliver talks, and attend top industry conferences.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now