遇见数据集

Salesforce/ContextualBench

收藏
Hugging Face2025-01-10 更新2025-04-08 收录
官方服务:

资源简介:

ContextualBench是一个强大的评估框架,旨在评估大型语言模型在上下文数据集上的性能。它提供了灵活的管道,用于评估各种LLM家族在不同任务上的表现,特别关注处理大型上下文输入。

ContextualBench is a powerful evaluation framework designed to assess the performance of Large Language Models (LLMs) on contextual datasets. It provides a flexible pipeline for evaluating various LLM families across different tasks, with a focus on handling large context inputs.

提供机构:
Salesforce
二维码
社区交流群
二维码
科研交流群
商业服务