官方服务:
资源简介:
Preprocessed into 80ms frames
应用场景:
创建时间:
2021-01-06
相关数据集
homebrewltd/raw-speech-whispervq-v1
该数据集包含超过240万条英语自动语音识别(ASR)样本,使用了特定的训练集和令牌化方法。数据集的特征包括索引、令牌和文本,主要用于自动语音识别任务,语言为英语。数据集的大小类别为1M到10M之间,许可证为MIT。
Hugging Face2024-08-19 更新130
ylacombe/accents_hidden_states
--- dataset_info: features: - name: labels dtype: string - name: length dtype: int64 - name: labels_id dtype: int64 - name: hidden_states sequence: float32 splits: - name
Hugging Face2024-04-04 更新80
Inferred parameter summary.
Statistics of inferred parameters and a derived measure of multiplicative noise intensity. Base-10 logarithms are used for normality. MISSD is median intra-subject SD, and MITSD is median intra-trial
NIAID Data Ecosystem50
ADLSET
ADLSET contains a collection of audio-4D landmark pairs captured from one specific male subject. It contains about 80𝑚𝑖𝑛 news audio spoken in Mandarin. The news from the THUCNews dataset includes s
Mendeley Data2024-05-22 更新90
Audiototext
Audio to text model with tokenizer and predication it will help to achieve
kaggle2023-08-19 更新60



