ToyADMOS dataset
收藏资源简介:
ToyADMOS dataset is a machine operating sounds dataset of approximately 540 hours of normal machine operating sounds and over 12,000 samples of anomalous sounds collected with four microphones at a 48kHz sampling rate, prepared by Yuma Koizumi and members in NTT Media Intelligence Laboratories. The dataset consists of three sub-dataset: "toy car" for product inspection task, "toy conveyor" for fault diagnosis for fixed machine task, and "toy train" for fault diagnosis for moving machine task. Since the total size of the ToyADMOS dataset is over 440GB, each sub-dataset is split into 7-9 files by 7-zip (7z-format). The total size of the compressed dataset is approximately 180GB, and that of each sub-dataset is approximately 60GB. Download the zip files corresponding to sub-datasets of interest and use your favorite compression tool to unzip these split zip files. The detail of the dataset is described in [1] and GitHub: https://github.com/YumaKoizumi/ToyADMOS-dataset License: see the file named LICENSE.pdf [1] Yuma Koizumi, Shoichiro Saito, Noboru Harada, Hisashi Uematsu and Keisuke Imoto, "ToyADMOS: A Dataset of Miniature-Machine Operating Sounds for Anomalous Sound Detection," in Proc of Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA), 2019.
ToyADMOS数据集(ToyADMOS Dataset)是由NTT媒体智能实验室的小泉悠马及其团队成员构建的机器运行音频数据集。该数据集通过4支麦克风以48kHz采样率采集了约540小时的正常机器运行音频,以及超过12000条异常声音样本。本数据集包含三个子数据集:分别为面向产品检测任务的“玩具车(toy car)”、面向固定机器故障诊断任务的“玩具传送带(toy conveyor)”,以及面向移动机器故障诊断任务的“玩具火车(toy train)”。由于该数据集总容量超过440GB,每个子数据集已通过7-zip(7z格式)拆分为7至9个分卷压缩文件。压缩后的总数据集大小约为180GB,单个子数据集的大小约为60GB。用户可下载所需子数据集对应的分卷压缩包,使用任意解压工具完成拆分压缩包的解压操作。数据集详细信息可参阅文献[1]及GitHub仓库:https://github.com/YumaKoizumi/ToyADMOS-dataset。许可证相关信息请查看名为LICENSE.pdf的文件。[1] 小泉悠马、斋藤翔一、原田升、植松久、井本启佑:“ToyADMOS:面向异常声音检测的微型机器运行声音数据集”,收录于2019年IEEE音频与声学信号处理应用研讨会(WASPAA)论文集。




