遇见数据集

inceptionai/Arabic_IFEval

收藏
Hugging Face2025-04-05 更新2025-04-08 收录
官方服务:

资源简介:

IFEval是一个首个公开的基准数据集,专门设计用于评估阿拉伯大型语言模型在阿拉伯语指令遵循能力上的表现。该数据集包括404个高质量的手动验证样本,这些样本涵盖了语言模式、标点符号规则和格式化指南等不同的限制条件。

IFEval is the first publicly available benchmark dataset specifically designed to evaluate Arabic Large Language Models (LLMs) on instruction-following capabilities in Arabic. The dataset includes 404 high-quality, manually verified samples covering various constraints such as linguistic patterns, punctuation rules, and formatting guidelines.

提供机构:
inceptionai
二维码
社区交流群
二维码
科研交流群
商业服务