遇见数据集

Profiles of rankings by the Mallows model

收藏
Zenodo2025-10-24 更新2026-05-26 收录
官方服务:

资源简介:

Dataset description This dataset contains profiles of rankings sampled using the Mallows model (Mallows, 1957), through the Repeated Insertion Model (Doignon et al, 2004). Profiles with different characteristics are included, and the following parameter values were considered: Number of alternatives: {3, 4, …, 16, 20, 25, 30} Number of voters: {10, 25, 75, 100, 125, …, 1000} Mallows model dispersion: {0.1, 0.4, 0.5, 0.6, 0.7, 0.75, 0.8, 0.85, 0.9, 0.95, 0.99, 0.999, 1.0} For every combination of the following parameters, there are 1000 profiles without nomalising the dispersion and 1000 profiles with the normalised dispersion (Boehmer et al, 2023) regarding the number of alternatives. The central ranking considered can be set when loading the data from disk. If you use this dataset, please cite the Zenodo entry. Central ranking considerations We have design this dataset trying to maximise its flexibility. Thus, we have sampled permutations of the central ranking considered for the Mallows model. When loading a profile, these permutations are mapped with the central ranking to obtain the final rankings. By doing so, the user can choose which central ranking to consider, rather than sampling the data again for that central ranking. For example, the following profile for 4 alternatives, given by: Permutation Votes [1, 2, 0, 3] 3 [0, 3, 2, 1] 2 is transformed to the final profile by mapping the permutation of the central ranking for each ranking of the profile. For instance, for the central ranking [3, 2, 1, 0], the final profile would be: Ranking Votes [2, 1, 3, 0] 3 [3, 0, 1, 2] 2 File distribution The dataset is provided as two zip files, as it exceeded the number of files limit. Each zip file gathers the profiles corresponding to normalising or not the dispersion. Please, decompress them in the directory used for storing the dataset locally. The linked GitHub repository gathers the code used for the generation of the data, and includes the functions to load the profiles and calculate the outranking matrix of a profile. Summary of changes to previous version The dataset has been updated to be suitable for use regardless of the central rankings considered by the authors. This has been achieved by sampling permutations of the central ranking. The structure of the Parquet dataset has been modified for greater efficiency, enabling users to download data with or without dispersion normalisation independently. Data for dispersion levels of 0.99 and 0.999 has also been added.

提供机构:
Zenodo
创建时间:
2025-05-02
二维码
社区交流群
二维码
科研交流群
商业服务