Danielbrdz/Barcenas-De-Cervantes
收藏资源简介:
--- license: cc-by-4.0 language: - es size_categories: - 10K<n<100K --- Barcenas de Cervantes Dataset con 50,739 ejemplos obtenidos por CORPES de la RAE Contiene textos en español escritas por Humanos de México, España, Argentina, Chile, Colombia, etc Adaptados con una teoría mía de conversación con el objetivo de mejorar el español de LLM pequeños, que son modelos que sufren mucho en el idioma de cervantes Muchas gracias a la RAE y a todos los autores de los escritos de los diferentes paises hispanos --------------------------------------------- Barcenas de Cervantes Dataset with 50,739 examples obtained from the RAE's CORPES system Contains Spanish texts written by people from Mexico, Spain, Argentina, Chile, Colombia, etc. Adapted with my own conversational theory to improve the Spanish of young LLMs, who are models that struggle a lot with the language of Cervantes. Many thanks to the RAE and all the authors of the texts from the different Spanish-speaking countries. Made with ❤️ in Guadalupe, Nuevo Leon, Mexico 🇲🇽



