Jingju a cappella singing dataset part1
收藏资源简介:
This is the 4th version of the dataset. The folder structure has been changed since the 2nd version, where the Laosheng folder has been moved directly into wav or textgrid folder. Description: This dataset is a collection of boundary annotations of a cappella singing performed by Beijing Opera (Jingju, 京剧) professional and amateur singers. wav.zip: audio files in .wav format, mono or stereo. wav_mono.zip: audio files in .wav format, mono pycode.zip: util code for parsing the .textgrid annotation catalogue*.csv: recording metadata, source separation recordings are not included. textgrid.zip: phrase, syllable and phoneme annotation in Praat .textgrid format annotation_txt.zip: phrase, syllable and phoneme time boundaries (second) and labels in .txt format *phrase_char: phrase-level time boundaries, labeled in Mandarin characters *phrase: phrase-level time boundaries, labeled in Mandarin pinyin *syllable: syllable-level time boundaries, labeled in Mandarin pinyin *phoneme: phoneme-level time boundaries, labeled in X-SAMPA The boundaries (onset and offset) have been annotated in both Praat TextGrid (textgrid.zip) and .txt (annotation_txt.zip) format hierarchically: phrase (line), syllable, phoneme Singing units in pinyin and X-SAMPA have been annotated to a jingju a cappella singing audio dataset. The corresponding audio files are the a cappella singing arias recordings, which are stereo or mono, sampled at 44.1 kHz, and stored as .wav files. The .wav files are recorded by two institutes: those file names ending with ‘qm’ are recorded by C4DM, Queen Mary University of London; others file names ending with ‘upf’ or ‘lon’ are recorded by MTG-UPF. Additionally, another collection of 15 clean singing recordings is included in this dataset. They are extracted from the commercial recordings which originally contains karaoke accompaniment and mixed versions. If you use this audio dataset in your work, please cite (1) this dataset as well (2) the following publication: D. A. A. Black, M. Li, and M. Tian, “Automatic Identification of Emotional Cues in Chinese Opera Singing,” in 13th Int. Conf. on Music Perception and Cognition (ICMPC-2014), 2014, pp. 250–255. Details: Annotation format, units, parsing code and other information please refer to https://github.com/MTG/jingjuPhonemeAnnotation License: Textgrid annotations are licensed under Creative Commons Attribution-NonCommercial 4.0 International License. Wav audio ending with ‘upf’ or ‘lon’ is licensed under Creative Commons Attribution-NonCommercial 4.0 International. For the license of .wav audio ending with ‘qm’ from C4DM Queen Mary University of London, please refer to this page http://isophonics.org/SingingVoiceDataset Contact information: Rong Gong: rong<dot>gong<at>upf<dot>edu Rafael Caro Repetto: rafael<dot>caro<at>upf<dot>edu
本数据集为第4版。自第2版起,数据集的文件夹结构已更新,老生(Laosheng)文件夹已直接移入wav或TextGrid文件夹中。 本数据集收录了京剧(Jingju)专业与业余演唱者的无伴奏歌唱边界标注。 wav.zip:包含单声道或立体声的.wav格式音频文件。 wav_mono.zip:包含单声道.wav格式音频文件。 pycode.zip:用于解析.textgrid格式标注的实用代码。 catalogue*.csv:包含录音元数据,不包含源分离录音。 textgrid.zip:采用Praat TextGrid格式存储的乐句、音节与音素标注。 annotation_txt.zip:采用.txt格式存储的乐句、音节与音素的时间边界(单位:秒)与标签。 *phrase_char:以中文字符标注的乐句级时间边界。 *phrase:以汉语拼音标注的乐句级时间边界。 *syllable:以汉语拼音标注的音节级时间边界。 *phoneme:以X-SAMPA音标标注的音素级时间边界。 边界(起始点与终止点)采用分层结构分别以Praat TextGrid(textgrid.zip)与.txt(annotation_txt.zip)格式进行标注,层级依次为乐句(行级)、音节与音素。 本京剧无伴奏歌唱音频数据集已完成基于汉语拼音与X-SAMPA音标的歌唱单元标注。对应音频文件为无伴奏歌唱选段录音,包含单声道与立体声两种格式,采样率为44.1 kHz,存储为.wav格式文件。 所有.wav音频文件由两家机构录制:文件名以「qm」结尾的文件由伦敦大学玛丽女王学院C4DM(Centre for Digital Music, Queen Mary University of London)录制;文件名以「upf」或「lon」结尾的文件由MTG-UPF录制。 此外,本数据集还收录了15段纯净歌唱录音,这些录音源自原本包含卡拉OK伴奏与混合版本的商业录音。 若您在研究工作中使用本音频数据集,请引用以下两项内容:(1) 本数据集;(2) 下述论文:D. A. A. Black、M. Li与M. Tian,《Automatic Identification of Emotional Cues in Chinese Opera Singing》,发表于第13届国际音乐感知与认知会议(ICMPC-2014),2014年,第250–255页。 详细信息(包括标注格式、标注单元、解析代码等)请参阅:https://github.com/MTG/jingjuPhonemeAnnotation 许可协议: TextGrid标注采用知识共享署名-非商业性使用4.0国际许可协议(CC BY-NC 4.0)进行授权。 文件名以「upf」或「lon」结尾的.wav音频文件采用知识共享署名-非商业性使用4.0国际许可协议进行授权。 文件名以「qm」结尾的.wav音频文件(由伦敦大学玛丽女王学院C4DM录制)的授权信息,请参阅:http://isophonics.org/SingingVoiceDataset 联系方式: 荣龚(Rong Gong):rong<dot>gong<at>upf<dot>edu 拉斐尔·卡雷诺·雷佩托(Rafael Caro Repetto):rafael<dot>caro<at>upf<dot>edu




