Carnatic Varnam Dataset
收藏资源简介:
Carnatic varnam dataset is a collection of 28 solo vocal recordings, recorded for our research on intonation analysis of Carnatic raagas. The collection has the audio recordings, taala cycle annotations and notations in a machine readable format. <strong>Audio music content</strong> They feature 7 varnams in 7 rāgas sung by 5 young professional singers who received training for more than 15 years. They are all set to Adi taala. Measuring the intonation variations require absolutely clean pitch contours. For this, all the varṇaṁs are recorded without accompanying instruments, except the drone. <strong>Taala annotations</strong> The recordings are annotated with taala cycles, each annotation marking the starting of a cycle. We have later automatically divided each cycle into 8 equal parts. The annotations are made available as sonic visualizer annotation layers. Each annotation is of the format m.n where m is the cycle number and n is the division within the cycle. All m.1 annotations are manually done, whereas m.[2-8] are automatically labelled. <strong>Notations</strong> The notations for 7 varnams are procured from an archive curated by Shivkumar, in word document format. They are manually converted to a machine readable format (yaml). Each file is essentially a dictionary with section names of the composition as keys. Each section is represented as a list of cycles. Each cycle in turn has a list of divisions. <strong>Notations</strong> The notation is given a single time per section, however, to align the svaras with the tala annotations, structure information is given. The structure is given in yaml format, specifying the order of the sections, and how many svaras are sung per each tala tick. Broadly, there are just two only cases, 2 svaras per tick, and 4 svaras per tick.<br> The structure information has been added in the 1.1 version of the dataset. No code is given to load the structure information and relate it with the rest of data in the dataset. You may refer to the mirdata loader of the Carnatic Varnam dataset where in this case, tools to easily load the dataset are given. The structure annotations and the mirdata loader have been curated and implemented by Adithi Shankar and Genís Plaja. <strong>Possible uses of the dataset</strong> The distinct advantage of this dataset is the free availability of the audio content. Along with the annotations, it can be used for melodic analyses: characterizing intonation, motif discovery and tonic identification. The availability of a machine readable notation files allows the dataset to be used for audio-score alignment. Using this dataset If you use this dataset in a publication, please cite: Koduri, G. K., Ishwar, V., Serrà, J., & Serra, X. (2014). Intonation analysis of rāgas in Carnatic music. Journal of New Music Research, 43(01), 73–94. http://hdl.handle.net/10230/25676 We are interested in knowing if you find our datasets useful! If you use our dataset please email us at mtg-info@upf.edu and tell us about your research. <strong>Contact</strong> If you have any questions or comments about the dataset, please feel free to write to us. <br> Music Technology Group,<br> Universitat Pompeu Fabra,<br> Barcelona, Spain<br> mtg-info <at> upf <dot>edu http://compmusic.upf.edu/carnatic-varnam-dataset
卡纳提克瓦尔纳姆(Carnatic varnam)数据集是一个包含28段独唱人声录音的合集,专为我们开展卡纳提克拉格(Carnatic rāgas)音调分析研究录制。该合集包含音频录音、塔拉(taala)循环标注以及机器可读格式的乐谱。 <strong>音频内容</strong> 本次录音涵盖7首瓦尔纳姆作品,分别对应7种拉格(rāgas),由5名拥有15年以上专业训练的青年职业歌手演唱,所有作品均采用阿迪塔拉(Adi taala)节拍体系。为精准测量音调变化,需获取纯净无干扰的音高轮廓,因此所有瓦尔纳姆录音仅搭配持续低音(drone)伴奏,未添加其他伴奏乐器。 <strong>塔拉循环标注</strong> 录音已完成塔拉循环标注,每一处标注均标记一个循环的起始位置。后续我们通过自动化手段将每个循环均分为8个均等的节拍分段。标注以声波可视化器(sonic visualizer)标注层的形式提供,每条标注采用`m.n`格式:其中`m`代表循环编号,`n`代表该循环内的分段编号。所有`m.1`类标注均为人工完成,而`m.[2-8]`类标注则为自动化标注。 <strong>乐谱标注</strong> 7首瓦尔纳姆的乐谱源自Shivkumar整理的档案库,初始格式为Word文档。我们已将其人工转换为机器可读的YAML格式。每份乐谱文件本质为一个字典,以作品的段落名称作为键值;每个段落由若干循环列表构成,而每个循环又包含若干分段列表。 <strong>乐谱结构标注</strong> 每个段落仅提供一次基础乐谱标注,但为实现斯瓦拉(svaras)与塔拉标注的对齐,我们额外提供了结构信息。该结构信息以YAML格式呈现,用于指定段落的排列顺序,以及每个塔拉节拍所演唱的斯瓦拉数量。总体而言仅存在两种场景:每个节拍对应2个斯瓦拉,或每个节拍对应4个斯瓦拉。 该结构信息已在数据集的1.1版本中加入。目前未提供用于加载结构信息并将其与数据集其余内容关联的代码。您可参考卡纳提克瓦尔纳姆数据集的mirdata加载工具,该工具可便捷实现数据集的加载。上述结构标注及mirdata加载工具由Adithi Shankar与Genís Plaja整理并实现。 <strong>数据集适用场景</strong> 本数据集的显著优势在于其音频内容可免费获取。结合配套标注,该数据集可用于旋律分析相关研究:包括音调特征刻画、动机发现以及主音识别。机器可读格式的乐谱文件,使得该数据集可应用于音频-乐谱对齐任务。 <strong>数据集使用规范</strong> 若您在学术出版物中使用本数据集,请引用以下文献:Koduri, G. K., Ishwar, V., Serrà, J., & Serra, X. (2014). Intonation analysis of rāgas in Carnatic music. Journal of New Music Research, 43(01), 73–94. http://hdl.handle.net/10230/25676 我们十分期待了解您是否认为本数据集具有实用价值!若您使用了本数据集,请发送邮件至mtg-info@upf.edu告知我们您的研究方向。 <strong>联系方式</strong> 若您对本数据集有任何疑问或建议,欢迎随时与我们联系。<br>音乐技术组(Music Technology Group)<br>庞培法布拉大学(Universitat Pompeu Fabra)<br>西班牙巴塞罗那<br>mtg-info <at> upf <dot>edu<br>http://compmusic.upf.edu/carnatic-varnam-dataset



