Spotify · via Lever

Senior Applied Research Scientist - Personalization

New York, NY · Remote

Responsibilities

  • Develop and experiment with new methods for speech synthesis and speech recognition, along with end-to-end approaches, building on the latest research and ideas.
  • Work towards the expansion of our speech use-cases targeting different markets and products.
  • Be part of a highly motivated research team dedicated to building and creating models at scale to power the Spotify platform.
  • Champion best practices for research and development, sharing your knowledge and experience with other researchers within Speak.
  • Collaborate with our engineering and data teams on ideas requiring new infrastructure or new high-quality data, as well as to help improve our speech recognition and speech synthesis pipelines, and help turn proven ideas into scalable products.

Requirements

  • You have a strong background in ML (PhD degree on top of professional experience), and
  • experience in working with any of the following: transformers, GANs, diffusion models, flow matching, VAEs, audio codecs.
  • You have experience in developing generative models for speech synthesis, speech recognition, audio/music, natural language processing, or computer vision.
  • You have strong experience with Python, particularly PyTorch.
  • You have strong communication skills and the ability to explain technical ideas with clarity to technical and non-technical people alike.
  • You have experience in an academic or professional setting conducting high-quality research.
  • This role is based in New York City.
  • We offer you the flexibility to work where you work best! There will be some in person meetings, but still allows for flexibility to work from home