Salesforce/ContextualBench
收藏官方服务:
资源简介:
ContextualBench是一个强大的评估框架,旨在评估大型语言模型在上下文数据集上的性能。它提供了灵活的管道,用于评估各种LLM家族在不同任务上的表现,特别关注处理大型上下文输入。
ContextualBench is a powerful evaluation framework designed to assess the performance of Large Language Models (LLMs) on contextual datasets. It provides a flexible pipeline for evaluating various LLM families across different tasks, with a focus on handling large context inputs.
提供机构:
Salesforce


