GoTriple Language Identification Benchmark Datasets
收藏官方服务:
资源简介:
This is a collection of two datasets used to benchmark various Language identification techniques both on the titles and abstracts of documents. They consist of two columns: Text: the text contained in the string Real Language: the language(s) contained in the string
提供机构:
Zenodo创建时间:
2025-11-04



