CatMeows: A Publicly-Available Dataset of Cat Vocalizations
收藏资源简介:
Abstract This dataset, composed of 440 sounds, contains meows emitted by cats in different contexts. Specifically, 21 cats belonging to 2 breeds (Maine Coon and European Shorthair) have been repeatedly exposed to three different stimuli that were expected to induce the emission of meows: Brushing - Cats were brushed by their owners in their home environment for a maximum of 5 minutes; Isolation in an unfamiliar environment - Cats were transferred by their owners into an unfamiliar environment (e.g., a room in a different apartment or an office). Distance was minimized and the usual transportation routine was adopted so as to avoid discomfort to animals. The journey lasted less than 30 minutes and cats were allowed 30 minutes with their owners to recover from transportation, before being isolated in the unfamiliar environment, where they stayed alone for maximum 5 minutes; Waiting for food - The owner started the routine operations that preceded food delivery in the usual environment the cat was familiar with. Food was given at most 5 minutes after the beginning of the experiment. The dataset has been produced and employed in the context of an interdepartmental project of the University of Milan (for further information, please refer to this doi). The content of the dataset has been described in detail in a scientific work currently under review; the reference will be provided as soon as the paper is published. File naming conventions Files containing meows are in the dataset.zip archive. They are PCM streams (.wav). Naming conventions follow the pattern C_NNNNN_BB_SS_OOOOO_RXX, which has to be exploded as follows: C = emission context (values: B = brushing; F = waiting for food; I: isolation in an unfamiliar environment); NNNNN = cat’s unique ID; BB = breed (values: MC = Maine Coon; EU: European Shorthair); SS = sex (values: FI = female, intact; FN: female, neutered; MI: male, intact; MN: male, neutered); OOOOO = cat owner’s unique ID; R = recording session (values: 1, 2 or 3) XX = vocalization counter (values: 01..99) Extra content The extra.zip archive contains excluded recordings (sounds other than meows emitted by cats) and uncut sequences of close vocalizations. Terms of use The dataset is open access for scientific research and non-commercial purposes. The authors require to acknowledge their work and, in case of scientific publication, to cite the most suitable reference among the following entries: Ntalampiras, S., Ludovico, L.A., Presti, G., Prato Previde, E., Battini, M., Cannas, S., Palestrini, C., Mattiello, S.: Automatic Classification of Cat Vocalizations Emitted in Different Contexts. Animals, vol. 9(8), pp. 543.1–543.14. MDPI (2019). ISSN: 2076-2615 Ludovico, L.A., Ntalampiras, S., Presti, G., Cannas, S., Battini, M., Mattiello, S.: CatMeows: A Publicly-Available Dataset of Cat Vocalizations. In: Li, X., Lokoč, J., Mezaris, V., Patras, I., Schoeffmann, K., Skopal, T., Vrochidis, S. (eds.) MultiMedia Modeling. 27th International Conference, MMM 2021, Prague, Czech Republic, June 22–24, 2021, Proceedings, Part II, LNCS, vol. 12573, pp. 230–243. Springer International Publishing, Cham (2021). ISBN: 978-3-030-67834-0 (print), 978-3-030-67835-7 (online) ISSN: 0302-9743 (print), 1611-3349 (online)
摘要 本数据集共包含440段音频素材,均为不同场景下猫咪发出的喵叫声。具体而言,来自缅因猫(Maine Coon)和欧洲短毛猫(European Shorthair)2个品种的21只猫咪,被先后施加三种预设可诱发喵叫的刺激: 1. 梳毛场景:猫咪在居家环境中由主人为其梳毛,单次时长不超过5分钟; 2. 陌生环境隔离:主人将猫咪转移至陌生环境(例如另一公寓的房间或办公室)。为避免动物不适,需尽量缩短运输距离并遵循常规运输流程,全程耗时不超过30分钟。待猫咪与主人共处30分钟以适应运输后,将其单独置于陌生环境中隔离,最长隔离时长为5分钟; 3. 待食场景:在猫咪熟悉的居家环境中,主人开始进行喂食前的常规操作,实验开始后最长5分钟内完成食物投喂。 本数据集由米兰大学跨部门项目制作并应用(更多信息请参见该DOI)。其内容已在一篇待发表的学术论文中详细阐述,论文正式发表后将提供具体引用信息。 文件命名规则 数据集压缩包dataset.zip中收录的喵叫音频均为脉冲编码调制(PCM)流格式的WAV音频文件。文件命名遵循`C_NNNNN_BB_SS_OOOOO_RXX`模式,各字段含义如下: - C:发声场景(取值:B=梳毛;F=待食;I=陌生环境隔离); - NNNNN:猫咪唯一编号; - BB:品种(取值:MC=缅因猫(Maine Coon);EU=欧洲短毛猫(European Shorthair)); - SS:性别(取值:FI=未绝育雌性;FN=已绝育雌性;MI=未绝育雄性;MN=已绝育雄性); - OOOOO:猫咪主人唯一编号; - R:录制场次(取值:1、2或3); - XX:发声计数(取值:01至99)。 额外内容 额外压缩包extra.zip中包含两类素材:一是被排除的录音(非猫咪喵叫的其他音频),二是未剪辑的连续喵叫序列。 使用条款 本数据集可开放获取用于科学研究及非商业用途。作者要求使用者在相关研究工作中注明本数据集的贡献,若用于学术发表,请从以下两篇文献中选择合适条目进行引用: 1. Ntalampiras, S., Ludovico, L.A., Presti, G., Prato Previde, E., Battini, M., Cannas, S., Palestrini, C., Mattiello, S.: Automatic Classification of Cat Vocalizations Emitted in Different Contexts. Animals, vol. 9(8), pp. 543.1–543.14. MDPI (2019). ISSN: 2076-2615 【译文】Ntalampiras S, Ludovico L A, Presti G, 等. 不同场景下猫咪喵叫声的自动分类[J]. 《Animals》, 2019, 9(8): 543.1-543.14. MDPI出版社。ISSN:2076-2615 2. Ludovico, L.A., Ntalampiras, S., Presti, G., Cannas, S., Battini, M., Mattiello, S.: CatMeows: A Publicly-Available Dataset of Cat Vocalizations. In: Li, X., Lokoč, J., Mezaris, V., Patras, I., Schoeffmann, K., Skopal, T., Vrochidis, S. (eds.) MultiMedia Modeling. 27th International Conference, MMM 2021, Prague, Czech Republic, June 22–24, 2021, Proceedings, Part II, LNCS, vol. 12573, pp. 230–243. Springer International Publishing, Cham (2021). ISBN: 978-3-030-67834-0 (print), 978-3-030-67835-7 (online) ISSN: 0302-9743 (print), 1611-3349 (online) 【译文】Ludovico L A, Ntalampiras S, Presti G, 等. CatMeows:公开可获取的猫咪喵叫数据集[C]//Li X, Lokoč J, Mezaris V, 等编. 第27届国际多媒体建模会议(MMM 2021)论文集(第二卷). 捷克布拉格, 2021年6月22-24日. 计算机科学论文集(LNCS)第12573卷. 沙姆:施普林格国际出版社, 2021: 230-243. ISBN:978-3-030-67834-0(印刷版)、978-3-030-67835-7(电子版);ISSN:0302-9743(印刷版)、1611-3349(电子版)




