遇见数据集

Danielbrdz/Barcenas-De-Cervantes

收藏
Hugging Face2026-04-09 更新2026-04-12 收录
官方服务:

资源简介:

--- license: cc-by-4.0 language: - es size_categories: - 10K<n<100K --- Barcenas de Cervantes Dataset con 50,739 ejemplos obtenidos por CORPES de la RAE Contiene textos en español escritas por Humanos de México, España, Argentina, Chile, Colombia, etc Adaptados con una teoría mía de conversación con el objetivo de mejorar el español de LLM pequeños, que son modelos que sufren mucho en el idioma de cervantes Muchas gracias a la RAE y a todos los autores de los escritos de los diferentes paises hispanos --------------------------------------------- Barcenas de Cervantes Dataset with 50,739 examples obtained from the RAE's CORPES system Contains Spanish texts written by people from Mexico, Spain, Argentina, Chile, Colombia, etc. Adapted with my own conversational theory to improve the Spanish of young LLMs, who are models that struggle a lot with the language of Cervantes. Many thanks to the RAE and all the authors of the texts from the different Spanish-speaking countries. Made with ❤️ in Guadalupe, Nuevo Leon, Mexico 🇲🇽

提供机构:
Danielbrdz
二维码
社区交流群
二维码
科研交流群
商业服务