遇见数据集

MXM

收藏
arXiv2025-09-30 收录
官方服务:

资源简介:

该数据集是百万歌曲数据集的官方歌词集合,它采用了5000个词汇量。该数据集包含了210,519个训练数据和27,143个测试数据点,旨在进行话题建模任务。

This dataset is the official lyric collection of the Million Song Dataset, which has a vocabulary size of 5000. It contains 210,519 training data points and 27,143 test data points, and is designed for topic modeling tasks.

提供机构:
Million Song Dataset
二维码
社区交流群
二维码
科研交流群
商业服务