遇见数据集

botbotrobotics/aya_dataset_pt

收藏
Hugging Face2024-02-29 更新2025-08-09 收录
官方服务:

资源简介:

--- language: - pt pretty_name: Aya Dataset Portuguese tags: - aya - portuguese - legal - chemistry license: apache-2.0 size_categories: - 1K<n<10K --- CohereForAI [Aya Dataset](https://huggingface.co/datasets/CohereForAI/aya_dataset) filtrado para português (PT). **Aya Dataset Summary** The [Aya Dataset](https://huggingface.co/datasets/CohereForAI/aya_dataset) is a multilingual instruction fine-tuning dataset curated by an open-science community via Aya Annotation Platform from Cohere For AI. The dataset contains a total of 204k human-annotated prompt-completion pairs along with the demographics data of the annotators. This dataset can be used to train, finetune, and evaluate multilingual LLMs. Curated by: Contributors of Aya Open Science Intiative. Language(s): 65 languages (71 including dialects & scripts). License: Apache 2.0

提供机构:
botbotrobotics
二维码
社区交流群
二维码
科研交流群
商业服务