遇见数据集

ITALIC: An Italian Intent Classification Dataset

收藏
Zenodo2023-06-14 更新2026-04-07 收录
数据链接:
官方服务:

资源简介:

<strong>ITALIC: An Italian Intent Classification Dataset</strong> ITALIC is a dataset of Italian audio recordings and contains annotation for utterance transcripts and associated intents. The ITALIC dataset was created through a custom web platform, utilizing both native and non-native Italian speakers as participants. The participants were required to record themselves while reading a randomly sampled short text from the MASSIVE dataset. ITALIC dataset containing 16,521 audio recordings collected by 70 different volunteers. The dataset is composed of: <em>recordings</em>: a folder containing the audio recordings in <em>.wav</em> format. It contains all the recordings composing the data collection. <em>[CONFIG_NAME]_[SPLIT_NAME].json</em>: the files containing metadata used for generating the configuration proposed in the paper and their corresponding splits: <em>[CONFIG_NAME]</em> is the name of the configuration, e.g. <em>massive</em>, <em>hard_noisy</em>, or <em>hard_speaker</em>. For the description of the configurations, please refer to the paper. <em>[SPLIT_NAME]</em> is the name of the split, e.g. <em>train</em>, <em>validation</em>, or <em>test</em>. Each split is different for each configuration. The metadata files are in JSON format, with one sample per line. Each sample is a JSON object with the following fields: <em>id</em>: the unique identifier of the sample. <em>age</em>: the age of the speaker (self-reported) <em>gender</em>: the gender of the speaker (self-reported) <em>region</em>: the region of origin of the speaker (self-reported) <em>nationality</em>: the nationality of the speaker (self-reported) <em>lisp</em>: the presence of a lisp in the speaker (self-reported) <em>education</em>: the education level of the speaker (self-reported) <em>speaker_id</em>: the unique identifier of the speaker <em>environment</em>: the environment in which the recording was made (self-reported) <em>device</em>: the device used for recording (self-reported) <em>scenario</em>, <em>field</em>, <em>intent</em>: the information parsed from massive annotations and accompanying metadata. <em>utt</em>: the utterance to be spoken by the speaker. This information is also taken from massive. <strong>Important Note:</strong> <strong>By downloading and accessing the dataset, you agree not to attempt to determine the identity of speakers in the ITALIC dataset or to clone their voices.</strong> <strong>License</strong> The ITALIC dataset is released under the Creative Commons Attribution 4.0 International License. If you use the dataset in your work, please cite the ITALIC paper.

创建时间:
2023-06-14
二维码
社区交流群
二维码
科研交流群
商业服务