Indian Regional Music Dataset
收藏资源简介:
This dataset is a collection of mel-spectrogram features extracted from Indian regional music containing the following languages:<br> Hindi, Gujarati, Marathi, Konkani, Bengali, Oriya, Kashmiri, Assamese, Nepali, Konyak, Manipuri, Khasi & Jaintia, Tamil, Malayalam, Punjabi, Telugu, Kannada. Five recordings are collected for each language for four artists (2Male + 2Female) each. 2 artists out of 4 for each language are old veteran performers, and the remaining 2 are contemporary performers. Overall, the dataset includes 17 languages, 68 artists (34 Males and 34 Females). There are 340 recordings in the dataset, with a total duration of 29.3 hrs. Mel-spectrogram is extracted from a 1-second segment with a 1/2 second sliding window for each song. Extracted mel-spectrogram for each segment is annotated with language, location, local_song_index, global_song_index, language_id, location_id, artist_id, gender_id. _________________________________________________________________________________________________________ This project was funded under the grant number: ECR/2018/000204 by the Science & Engineering Research Board (SERB).
本数据集为从印度地区性音乐中提取的梅尔频谱图(mel-spectrogram)特征集合,涵盖以下语种:印地语、古吉拉特语、马拉地语、孔卡尼语、孟加拉语、奥里亚语、克什米尔语、阿萨姆语、尼泊尔语、克尤语、曼尼普尔语、卡西族与詹蒂亚族语、泰米尔语、马拉雅拉姆语、旁遮普语、泰卢固语、卡纳达语。 针对每种语言,分别为4位艺人(2名男性+2名女性)各录制5段音频。每类语言对应的4位艺人中,2位为资深老牌表演者,剩余2位为当代艺人。 整体而言,本数据集涵盖17种语言、68位艺人(男性34名、女性34名),共包含340段录音,总时长达29.3小时。 针对每首歌曲,以1秒片段结合0.5秒滑动窗口提取梅尔频谱图特征。每段提取得到的梅尔频谱图均标注有语言、地域、本地歌曲索引(local_song_index)、全局歌曲索引(global_song_index)、语言ID(language_id)、地域ID(location_id)、艺人ID(artist_id)以及性别ID(gender_id)。 本项目由科学与工程研究委员会(Science & Engineering Research Board, SERB)以编号ECR/2018/000204的科研资助项目支持。



