遇见数据集

Spoken Arabic Regional Archive (SARA)

收藏
Mendeley Data2024-03-27 更新2024-06-26 收录
官方服务:

资源简介:

SARA is a database, which consists of a set of spontaneous not pre-specified colloquial phrases in everyday life and life situations that are collected from media shows, episodes and films published on YouTube played by native speakers with three different Arabic dialects and accents. SARA's dialects are the Egyptian dialect (EGY), the Arabian Peninsula dialect (ARP) and the Levantine dialect (LEV). In addition, within each dialect group there are a number of different accents. Each utterance contains a one-speaker speech, the number of speakers within the dialect and the number of utterances by a speaker is unknown. This dataset can be used in speech, speaker, dialect and accent recognition applications. SARA dataset contains only adult speakers to avoid the improper pronunciation of the children that can affect the detection process. The dataset samples are variant in length from 3 to 7 seconds in order to verify the minimum time in which we can determine the speaker dialect or accent when the speaker speaks in free talk, which is the main scope of this research.

SARA数据集是一个语料库,其数据采集自YouTube平台发布的媒体节目、剧集与电影,收录了由使用三种不同阿拉伯方言与口音的母语使用者录制的、日常及生活场景下的非预设自发口语短语集合。SARA覆盖的方言包含埃及方言(Egyptian Dialect, EGY)、阿拉伯半岛方言(Arabian Peninsula Dialect, ARP)以及黎凡特方言(Levantine Dialect, LEV);此外,每个方言组内还包含多种细分口音。每条语音片段均为单发言人语音,但该方言组内的发言人总数以及单发言人的语音片段数量均未知。该数据集可应用于语音识别、发言人识别、方言识别以及口音识别相关任务。为避免儿童发音不规范干扰识别流程,SARA数据集仅收录成年发言人的语音数据。数据集样本的时长跨度为3至7秒,旨在验证自由会话场景下,识别发言人方言或口音所需的最短时长——这也是本研究的核心研究范畴。

创建时间:
2024-01-23
二维码
社区交流群
二维码
科研交流群
商业服务