遇见数据集

Playlist2vec: Spotify Million Playlist Dataset

收藏
NIAID Data Ecosystem2026-03-12 收录
数据链接:
官方服务:

资源简介:

This dataset was created using Spotify developer API. It consists of user-created as well as Spotify-curated playlists. The dataset consists of 1 million playlists, 3 million unique tracks, 3 million unique albums, and 1.3 million artists. The data is stored in a SQL database, with the primary entities being songs, albums, artists, and playlists. Each of the aforementioned entities are represented by unique IDs (Spotify URI). Data is stored into following tables: album artist track playlist track_artist1 track_playlist1 album | id | name | uri | id: Album ID as provided by Spotify name: Album Name as provided by Spotify uri: Album URI as provided by Spotify artist | id | name | uri | id: Artist ID as provided by Spotify name: Artist Name as provided by Spotify uri: Artist URI as provided by Spotify track | id | name | duration | popularity | explicit | preview_url | uri | album_id | id: Track ID as provided by Spotify name: Track Name as provided by Spotify duration: Track Duration (in milliseconds) as provided by Spotify popularity: Track Popularity as provided by Spotify explicit: Whether the track has explicit lyrics or not. (true or false) preview_url: A link to a 30 second preview (MP3 format) of the track. Can be null uri: Track Uri as provided by Spotify album_id: Album Id to which the track belongs playlist | id | name | followers | uri | total_tracks | id: Playlist ID as provided by Spotify name: Playlist Name as provided by Spotify followers: Playlist Followers as provided by Spotify uri: Playlist Uri as provided by Spotify total_tracks: Total number of tracks in the playlist. track_artist1 | track_id | artist_id | Track-Artist association table track_playlist1 | track_id | playlist_id | Track-Playlist association table - - - - - SETUP - - - - - The data is in the form of a SQL dump. The download size is about 10 GB, and the database populated from it comes out to about 35GB. spotifydbdumpschemashare.sql contains the schema for the database (for reference): spotifydbdumpshare.sql is the actual data dump. Setup steps: 1. Create database 2. mysql -u -p < spotifydbdumpshare.sql - - - - - PAPER - - - - - The description of this dataset can be found in the following paper: Papreja P., Venkateswara H., Panchanathan S. (2020) Representation, Exploration and Recommendation of Playlists. In: Cellier P., Driessens K. (eds) Machine Learning and Knowledge Discovery in Databases. ECML PKDD 2019. Communications in Computer and Information Science, vol 1168. Springer, Cham

创建时间:
2021-06-22
二维码
社区交流群
二维码
科研交流群
商业服务