遇见数据集

jknafou/TransCorpus-bio-hi

收藏
Hugging Face2025-12-03 更新2025-10-25 收录
官方服务:

资源简介:

TransCorpus-bio-hi是一个大规模的平行生物医学语料库,包含PubMed摘要的印地语合成翻译。该数据集使用TransCorpus框架和M2M-100模型创建,旨在支持高质量的印地语生物医学语言模型训练和下游NLP研究。

TransCorpus-bio-hi is a large-scale parallel biomedical corpus consisting of synthetic Hindi translations of PubMed abstracts. Created using the TransCorpus framework and the M2M-100 model, this dataset is designed to support high-quality Hindi biomedical language modeling and downstream NLP research.

提供机构:
jknafou
二维码
社区交流群
二维码
科研交流群
商业服务