zjs555888/DogSpeak_Dataset
收藏资源简介:
DogSpeak是一个大规模、野外采集的犬类发声数据集,旨在推动动物通信和计算生物声学的研究。与以往在受控环境中记录的数据集不同,DogSpeak来源于数万个社交媒体视频,捕捉了各种自然、有机的互动。该数据集包含77,202个犬吠序列(称为Barkseqs),来自156只不同个体的狗,涵盖5个品种:吉娃娃、德国牧羊犬、哈士奇、比特犬和柴犬。数据集组织为简单的目录结构,每个狗的音频剪辑位于以dog_id命名的文件夹中,所有音频文件均为.wav格式。由于存储限制,原始dog_7文件夹被拆分为两个子目录dog_7a和dog_7b。元数据文件metadata.csv提供了每个音频剪辑的关键信息,包括文件名、品种、性别和狗的唯一ID。
DogSpeak is a large-scale, in-the-wild canine vocalization dataset designed to advance research in animal communication and computational bioacoustics. Unlike previous datasets recorded in controlled environments, DogSpeak is sourced from tens of thousands of online social media videos, capturing a wide array of natural, organic interactions. The dataset contains 77,202 bark sequences (referred to as Barkseqs) from 156 individual dogs across 5 breeds: Chihuahua, German Shepherd, Husky, Pitbull, and Shiba Inu. The dataset is organized into a simple directory structure where each dogs audio clips are located in a folder named with a sequential dog_id. All audio files are in .wav format. Due to a repository limit, the original dog_7 folder was split into two subdirectories: dog_7a and dog_7b. The metadata.csv file provides key information for each audio clip, including filename, breed, sex, and dog_id.




