遇见数据集

ahmad21omar/SLR-Bench-Italian

收藏
Hugging Face2025-10-17 更新2025-10-25 收录
官方服务:

资源简介:

SLR-Bench-Italian是SLR-Bench的意大利语版本,保持了与英语版本相同的符号结构、评估框架和课程,但所有自然语言任务提示都翻译成了意大利语。数据集包含超过19,000个任务,每个任务都包含一个自然语言提示、一个可执行的验证程序以及一个潜在的地面真实规则。数据集被分为20个复杂度级别,分为4个广泛的层次(基础、简单、中等、困难)。数据集可以用来评估和训练大语言模型(LLMs)的逻辑推理能力,支持多语言推理和跨语言泛化研究。

SLR-Bench-Italian is the Italian-language pendant of the original SLR-Bench dataset. It follows the same symbolic structure, evaluation framework, and curriculum as the English version but provides all natural-language task prompts translated into Italian. This enables systematic evaluation and training of Large Language Models (LLMs) in logical reasoning in Italian, supporting both multilingual reasoning and cross-lingual generalization research.

提供机构:
ahmad21omar
二维码
社区交流群
二维码
科研交流群
商业服务