Data for: Bilinguals apply language-specific grain sizes during sentence reading
收藏官方服务:
资源简介:
Research data
研究数据
应用场景:
创建时间:
2024-01-23
相关数据集
yodas_owsmv4
该数据集包含了跨越75种语言的166,000小时的多种语言语音,被分割成30秒的长格式音频片段。数据来源于YODAS2数据集,该数据集基于大规模的网络爬取内容。由于网络源数据的性质,原始的YODAS2数据集可能包含不准确的语言标签和音频-文本对不齐的情况。为了解决这一问题,我们开发了一个可扩展的数据清洗管道,使用公开可用的工具包,从而形成原始数据集的一个精选子集。这个清洗后的数据集是我们OWSM
Hugging Face2025-06-03 更新390
Supplementary Material for: Differential Phonological and Semantic Modulation of Neurophysiological Responses to Visual Word Recognition
Background: Reading words for meaning relies on orthographic, phonological and semantic processing. The triangle model implicates a direct orthography-to-semantics pathway and a phonolog
DataCite Commons2020-09-02 更新120
ittailup/filtered_common_voice-hi-source
--- dataset_info: features: - name: audio dtype: audio: sampling_rate: 16000 - name: sentence dtype: string splits: - name: train num_bytes: 125144078.138 num_e
Hugging Face2024-05-12 更新90
Is the Lexical Boost Due to the Recency of the Repeated Word: Experimental Data, 2017-2022
In two structural priming experiments, participants read a Dutch prime sentence aloud, followed by a Dutch target fragment that they had to complete using pictures. Prime sentences were either double
CESSDA2025-06-12 更新130
techpurebroking/gujarati_dataset_ultravox_5
该数据集包含音频文件及其相关信息,特征字段包括音频ID、文件路径、音频数据、文本句子、上下投票数、年龄、性别、口音、地区、音频片段、变体和连续性标志。数据集分为训练集、验证集和测试集,各自包含不同的样本数量和大小。数据集总大小为2,076,252字节,下载大小为1,028,029字节。
Hugging Face2025-02-12 更新100



