相关数据集
julia-lukasiewicz-pater/GPT-wiki-intro-features
--- license: cc task_categories: - text-classification language: - en size_categories: - 100K<n<1M --- # Small-GPT-wiki-intro-features dataset This dataset is based on [aadityaubhat/GPT-wiki-intro](
Hugging Face2023-06-11 更新160
Replication Data for: Covering Blue Voices: African American English and Authenticity in Blues Covers
Repository Description This repository contains data for a quantitative analysis of blues lyrics performed by artists across time and socio-cultural groups. This analysis is a part of my PhD project o
DataONE2025-06-04 更新100
ASLP-lab/UrduSpeech
UrduSpeech是一个大规模、高保真的乌尔都语语音语料库,包含156小时的音频,并附有全面的12维副语言元数据。该语料库通过以下方式解决了乌尔都语在语音技术中资源严重不足的问题:- 71,792个经过对话者分割的话语,涵盖多样化的内容类别;- 三个专业子集:标准巴基斯坦乌尔都语(US-Std,59.2小时)、乌尔都语-英语语码转换(US-CS,89.4小时)和巴基斯坦口音英语(US-EngPk
Hugging Face2026-05-29 更新40
Additional file 1 of Predicting probable Alzheimerâs disease using linguistic deficits and biomarkers
Machine Learning files for all the models presented in our results, including baseline models. These files contain the transformed linguistic features from the DementiaBank dataset for both disease an
NIAID Data Ecosystem150
Heritage Japanese
This research was supported by the Center for Advanced Study of Language at the University of Maryland, the Faculty of Arts & Sciences at Harvard University, and the National Heritage Language Res
DataCite Commons2025-05-11 更新70



