DTU54DL/commonvoice_accent_test
收藏资源简介:
--- annotations_creators: - expert-generated language: - en language_creators: - found license: - mit multilinguality: - monolingual paperswithcode_id: acronym-identification pretty_name: Acronym Identification Dataset size_categories: - 10K<n<100K source_datasets: - original task_categories: - token-classification task_ids: - token-classification-other-acronym-identification train-eval-index: - col_mapping: labels: tags tokens: tokens config: default splits: eval_split: test task: token-classification task_id: entity_extraction --- # Dataset Card for [Dataset Name] ## Table of Contents - [Table of Contents](#table-of-contents) - [Dataset Description](#dataset-description) - [Dataset Summary](#dataset-summary) - [Supported Tasks and Leaderboards](#supported-tasks-and-leaderboards) - [Languages](#languages) - [Dataset Structure](#dataset-structure) - [Data Instances](#data-instances) - [Data Fields](#data-fields) - [Data Splits](#data-splits) - [Dataset Creation](#dataset-creation) - [Curation Rationale](#curation-rationale) - [Source Data](#source-data) - [Annotations](#annotations) - [Personal and Sensitive Information](#personal-and-sensitive-information) - [Considerations for Using the Data](#considerations-for-using-the-data) - [Social Impact of Dataset](#social-impact-of-dataset) - [Discussion of Biases](#discussion-of-biases) - [Other Known Limitations](#other-known-limitations) - [Additional Information](#additional-information) - [Dataset Curators](#dataset-curators) - [Licensing Information](#licensing-information) - [Citation Information](#citation-information) - [Contributions](#contributions) ## Dataset Description - **Homepage:** - **Repository:** - **Paper:** - **Leaderboard:** - **Point of Contact:** ### Dataset Summary [More Information Needed] ### Supported Tasks and Leaderboards [More Information Needed] ### Languages [More Information Needed] ## Dataset Structure ### Data Instances [More Information Needed] ### Data Fields [More Information Needed] ### Data Splits [More Information Needed] ## Dataset Creation ### Curation Rationale [More Information Needed] ### Source Data #### Initial Data Collection and Normalization [More Information Needed] #### Who are the source language producers? [More Information Needed] ### Annotations #### Annotation process [More Information Needed] #### Who are the annotators? [More Information Needed] ### Personal and Sensitive Information [More Information Needed] ## Considerations for Using the Data ### Social Impact of Dataset [More Information Needed] ### Discussion of Biases [More Information Needed] ### Other Known Limitations [More Information Needed] ## Additional Information ### Dataset Curators [More Information Needed] ### Licensing Information [More Information Needed] ### Citation Information [More Information Needed] ### Contributions Thanks to [@github-username](https://github.com/<github-username>) for adding this dataset.
--- annotations_creators: - 专家生成(expert-generated) language: - 英语(en) language_creators: - 现有文本采集(found) license: - MIT许可证(mit) multilinguality: - 单语言(monolingual) paperswithcode_id: acronym-identification pretty_name: 首字母缩写识别数据集(Acronym Identification Dataset) size_categories: - 10K<n<100K source_datasets: - 原始数据集(original) task_categories: - Token 分类(token-classification) task_ids: - Token 分类子类:首字母缩写识别(token-classification-other-acronym-identification) train-eval-index: - col_mapping: 标签: 标记 Token: Token config: 默认配置(default) splits: eval_split: 测试集(test) task: Token 分类(token-classification) task_id: 实体抽取(entity_extraction) --- ## [数据集名称] 数据集卡片 ## 目录 - [目录](#table-of-contents) - [数据集描述](#dataset-description) - [数据集摘要](#dataset-summary) - [支持任务与排行榜](#supported-tasks-and-leaderboards) - [语言](#languages) - [数据集结构](#dataset-structure) - [数据实例](#data-instances) - [数据字段](#data-fields) - [数据划分](#data-splits) - [数据集构建](#dataset-creation) - [构建依据](#curation-rationale) - [源数据](#source-data) - [标注信息](#annotations) - [个人与敏感信息](#personal-and-sensitive-information) - [数据使用注意事项](#considerations-for-using-the-data) - [数据集的社会影响](#social-impact-of-dataset) - [偏差讨论](#discussion-of-biases) - [其他已知局限性](#other-known-limitations) - [附加信息](#additional-information) - [数据集策展人](#dataset-curators) - [许可信息](#licensing-information) - [引用信息](#citation-information) - [贡献](#contributions) ## 数据集描述 - **主页:** - **代码仓库:** - **论文:** - **排行榜:** - **联系人:** ### 数据集摘要 [More Information Needed] ### 支持任务与排行榜 [More Information Needed] ### 语言 [More Information Needed] ## 数据集结构 ### 数据实例 [More Information Needed] ### 数据字段 [More Information Needed] ### 数据划分 [More Information Needed] ## 数据集构建 ### 构建依据 [More Information Needed] ### 源数据 #### 初始数据收集与归一化 [More Information Needed] #### 源语言生产者是谁? [More Information Needed] ### 标注信息 #### 标注流程 [More Information Needed] #### 标注者是谁? [More Information Needed] ### 个人与敏感信息 [More Information Needed] ## 数据使用注意事项 ### 数据集的社会影响 [More Information Needed] ### 偏差讨论 [More Information Needed] ### 其他已知局限性 [More Information Needed] ## 附加信息 ### 数据集策展人 [More Information Needed] ### 许可信息 [More Information Needed] ### 引用信息 [More Information Needed] ### 贡献 感谢 [@github-用户名](https://github.com/<github-用户名>) 添加此数据集。
数据集概述
- 名称: Acronym Identification Dataset
- 别名: 简称识别数据集
- 语言: 英语
- 语言创建者: 发现
- 许可证: MIT
- 多语言性: 单语种
- 大小: 10K<n<100K
- 来源: 原始数据
- 任务类别: 令牌分类
- 任务ID: token-classification-other-acronym-identification
- 训练与评估索引:
- 配置: 默认
- 分割:
- 评估分割: 测试
- 任务: 令牌分类
- 任务ID: 实体提取
数据集结构
- 数据实例: [未提供详细信息]
- 数据字段: [未提供详细信息]
- 数据分割: [未提供详细信息]
数据集创建
- 筛选理由: [未提供详细信息]
- 源数据: [未提供详细信息]
- 注释:
- 创建者: 专家生成
- 注释过程: [未提供详细信息]
- 注释者: [未提供详细信息]
- 个人和敏感信息: [未提供详细信息]
使用数据集的考虑
- 数据集的社会影响: [未提供详细信息]
- 偏见讨论: [未提供详细信息]
- 其他已知限制: [未提供详细信息]
附加信息
- 数据集管理员: [未提供详细信息]
- 许可信息: [未提供详细信息]
- 引用信息: [未提供详细信息]
- 贡献: 感谢@github-username添加此数据集。



