登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
Exemplified NMT, ST and CG warm-up training programs.
Exemplified NMT, ST and CG warm-up training programs.
收藏
Figshare
2025-02-21 更新
2026-04-28 收录
神经机器翻译优化
代码生成预训练
数据链接:
https://figshare.com/articles/dataset/Exemplified_NMT_ST_and_CG_warm-up_training_programs_/28461500
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
Exemplified NMT, ST and CG warm-up training programs.
应用场景:
创建时间:
2025-02-21
相关数据集
kp7742/YALM-pretrain3-122M
多语言文本生成
代码生成预训练
YALM预训练数据集-3是一个包含数学、Python代码和英语、印地语、古吉拉特语多语言数据的混合数据集,用于语言建模任务和YALM语言模型的开发。总样本数约为1.22亿,测试集样本数为2.2万。数据集未随机排序,而是连续拼接。数据集来源于多个子数据集,包括英语、印地语、古吉拉特语文本和Python代码。
Hugging Face
2025-04-02 更新
6
0
ACT-MNMT: Auto-Constriction Turning for Multilingual Neural Machine Translation
多语言机器翻译
神经机器翻译优化
Multilingual Neural Machine Translation
Mendeley Data
1
0
ACT-MNMT: Auto-Constriction Turning for Multilingual Neural Machine Translation
多语言机器翻译
神经机器翻译优化
Multilingual Neural Machine Translation
Mendeley Data
0
0
adorkin/nemotron-code-student-teacher-10M
代码生成预训练
合成代码数据
--- dataset_info: features: - name: text dtype: string - name: n_tokens dtype: int64 splits: - name: train num_bytes: 26474862563.0 num_examples: 10000000 download_size: 13
Hugging Face
2026-04-24 更新
3
0
REVIEWER PROCESSES AND PARALLEL DICTIONARY CREATION IN ANNOTATING UZBEK DIALECTS FOR TRAINING ARTIFICIAL INTELLIGENCE MODELS ON THE EXAMPLE OF THE KHOREZM DIALECT
低资源方言标注
神经机器翻译优化
This study presents the methodology for annotating the Khorezm dialect, a low-resource variety of the Uzbek language, and developing parallel dictionaries for training artificial intelligence (AI) mod
Zenodo
2026-06-03 更新
0
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广