遇见数据集

Catalan Sub-word Embeddings in FastText

收藏
Zenodo2021-02-25 更新2026-04-07 收录
数据链接:
官方服务:

资源简介:

These Catalan sub-word embeddings in FastText using BPE have been generated from the largest corpus ever made in Catalan till the date. The corpus has more than 10Gb of curated high quality text. If this material is useful, please cite it. Copyright (c) 2021 Text Mining Unit - Barcelona Supercomputing Center

创建时间:
2021-02-25
二维码
社区交流群
二维码
科研交流群
商业服务