lonelynode/elperuano-legal-corpus
收藏官方服务:
资源简介:
El Peruano Legal Corpus 是一个西班牙语法律文本数据集,专注于秘鲁政府相关文档。该数据集适用于文本分类和标记分类等自然语言处理任务,内容可能涉及法律条文、政府公告或其他法律材料。数据规模在100万到1000万之间,旨在支持法律AI应用、文本预处理和基准测试。标签包括法律、NLP、西班牙语、秘鲁、政府等,表明其用途广泛,可用于研究和开发。
El Peruano Legal Corpus is a Spanish legal text dataset focused on Peruvian government-related documents. It is suitable for natural language processing tasks such as text classification and token classification, and may include legal provisions, government announcements, or other legal materials. The dataset size ranges from 1 million to 10 million entries, designed to support legal AI applications, text preprocessing, and benchmarking. Tags include legal, NLP, Spanish, Peru, government, etc., indicating its broad utility for research and development.
提供机构:
lonelynode


