jknafou/TransCorpus-bio-hi
收藏官方服务:
资源简介:
TransCorpus-bio-hi是一个大规模的平行生物医学语料库,包含PubMed摘要的印地语合成翻译。该数据集使用TransCorpus框架和M2M-100模型创建,旨在支持高质量的印地语生物医学语言模型训练和下游NLP研究。
TransCorpus-bio-hi is a large-scale parallel biomedical corpus consisting of synthetic Hindi translations of PubMed abstracts. Created using the TransCorpus framework and the M2M-100 model, this dataset is designed to support high-quality Hindi biomedical language modeling and downstream NLP research.
提供机构:
jknafou


