Netflix Prize Data Set
收藏资源简介:
This is the official data set used in the Netflix Prize competition. The data consists of about 100 million movie ratings, and the goal is to predict missing entries in the movie-user rating matrix. |Attribute| Value| |——|—-| | Data Set Characteristics: | Multivariate, Time-Series | | Attribute Characteristics: | Integer | | Associated Tasks: | Clustering, Recommender-Systems | | Number of Instances: | 100480507 | | Number of Attributes: | 17770 | | Missing Values? | Yes | | Area: | N/A | #Data Set Information: This dataset was constructed to support participants in the Netflix Prize. There are over 480,000 customers in the dataset, each identified by a unique integer id. The title and release year for each movie is also provided. There are over 17,000 movies in the dataset, each identified by
本数据集为Netflix Prize竞赛的官方数据集。数据包含约一亿条电影评分记录,旨在预测电影与用户评分矩阵中的缺失项。 |属性|值| |——|—-| 数据集特征: |——|多变量、时间序列| 属性特征: |——|整数| 相关任务: |——|聚类、推荐系统| 实例数量: |——|100480507| 属性数量: |——|17770| 缺失值: |——|是| 区域: |——|N/A| #数据集信息: 本数据集旨在辅助Netflix Prize竞赛的参与者。数据集中包含超过48万用户,每位用户以唯一的整数ID进行标识。此外,每部电影的标题和发行年份也被提供。数据集中包含超过17,000部电影,每部电影以唯一的整数ID进行标识。




