登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
StephanAkkerman/english-words-IPA-embeddings
StephanAkkerman/english-words-IPA-embeddings
收藏
Hugging Face
2024-12-18 更新
2024-12-14 收录
自然语言处理
多语言学习
数据链接:
https://hf-mirror.com/datasets/StephanAkkerman/english-words-IPA-embeddings
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
--- license: mit ---
许可证:MIT许可证(MIT License)
应用场景:
提供机构:
StephanAkkerman
相关数据集
GENIAC-team-haijima/yutohub-002-sft-data
机器学习
自然语言处理
--- dataset_info: features: - name: output dtype: string - name: input dtype: string - name: instruction dtype: string splits: - name: train num_bytes: 8959599 num_exam
Hugging Face
2024-05-26 更新
17
0
aslawliet/A-Math
数学教育
自然语言处理
--- license: cc-by-nc-2.0 task_categories: - text-generation - question-answering language: - en ---
Hugging Face
2024-03-15 更新
10
0
Luganda Monolingual Corpus
自然语言处理
语言学
This dataset contains 100,000 Luganda sentences. Luganda is a Bantu language and is one of the major languages spoken in Uganda. This dataset was compiled by researchers at the Makerere AI and Data Sc
DataCite Commons
2025-04-15 更新
7
0
prli/pubmed-full_chopped_pyt_llama_alpha_0-1
医学文本处理
自然语言处理
该数据集包含文本、元信息、句子扰动数量以及文档统计信息。文本是数据集的主要特征,元信息中包含了数据集的集合名称。文档统计信息详细记录了文档的优势分数、长度变化比例、原始句子长度、替换后的句子长度等多个统计数据。数据集分为验证集,验证集大小为4000个样本。
Hugging Face
2025-10-14 更新
13
0
tyzhu/find_last_sent_train_100_eval_10_hint5
自然语言处理
文本分析
--- dataset_info: features: - name: inputs dtype: string - name: targets dtype: string - name: title dtype: string - name: context dtype: string splits: - name: train
Hugging Face
2023-10-31 更新
8
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广