Wbbyyr: FastText language models for Mandarin Chinese, trained on 14m Sina Weibo posts for each year in 2012-2018 (Fold 1 of 10)
收藏数据链接:
官方服务:
资源简介:
Wbbyyr: FastText language models for Mandarin Chinese, trained on 14,440,000 Sina Weibo posts for each year in 2012-2018. The 14,440,000 posts from each year are split into 10 folds. Due to Zenodo size limit, this dataset contains only the first fold from each year. Each model is trained for 20 iterations. Each vector is 300 dimensions long.
提供机构:
Zenodo创建时间:
2020-01-12



