Topic

Semantic IDs

All digests tagged Semantic IDs

Teaching LLMs to Speak Spotify — Yves Raimond & Jacqueline Wood, Spotify thumbnail

· 19:40

Teaching LLMs to Speak Spotify — Yves Raimond & Jacqueline Wood, Spotify

Spotify has transitioned its recommendation system from traditional curation and ranking algorithms to 'generative personalization' by making Large Language Models (LLMs) native to its catalog. The core technology, the Large Taste Model, allows users to interact with the system using natural language, enabling features like steerable DJs, prompted playlists, and editable taste profiles. The system is powered by a four-stage training recipe called NEO, which embeds open-weight LLMs with Spotify's catalog knowledge using Semantic IDs, ensuring both high performance and the retention of core language abilities.

Key takeaways

  1. Shift to Generative Personalization 3:57

    Spotify moved from 'personalization as guessing' (ranking algorithms) to 'personalization as reasoning,' allowing the system to introspect and generate experiences dynamically shaped around the user's intent. This also represents a shift from black-box algorithms to transparent, steerable systems. (2:37)

  2. The Large Taste Model (LTM) 10:52

    The LTM is the central system that combines prediction and reasoning, allowing users to shape and generate experiences in real time. As of today, about one in four US Premium subscribers interact with it daily. (6:52, 7:27)

  3. LLM Judge Grounding for Evaluation 17:32

    Traditional offline metrics are insufficient for generative systems. To evaluate performance, Spotify grounds LLM judges using textual user profiles (summarizing listening history) and actual behavioral signals. Grounding on ambiguous queries increased alignment with human preferences by 91%. (10:52)

Watch on YouTube Full article