登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
ROUGE evaluation on the SAMSum and DialogSum test sets.
ROUGE evaluation on the SAMSum and DialogSum test sets.
收藏
Figshare
2024-04-16 更新
2026-04-28 收录
对话评估
ROUGE评估指标
数据链接:
https://figshare.com/articles/dataset/ROUGE_evaluation_on_the_SAMSum_and_DialogSum_test_sets_/25613189
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
Results with * are obtained from [23],+ are obtained from [41].
应用场景:
创建时间:
2024-04-16
相关数据集
ScaleAI/MultiChallenge
对话评估
语言模型基准测试
--- license: cc-by-4.0 task_categories: - text-generation - question-answering language: - en tags: - multi-turn - evaluation - benchmark - llm pretty_name: MultiChallenge size_categories: - n<1K data
Hugging Face
2026-03-31 更新
17
0
PPL and BLEU scores with weighted discriminator and MLE loss for the topical chat frequent test set.
对话评估
自然语言生成
PPL and BLEU scores with weighted discriminator and MLE loss for the topical chat frequent test set.
NIAID Data Ecosystem
5
0
wyzard-ai/Trisha
CRM客户关系管理
对话评估
--- size_categories: n<1K tags: - rlfh - argilla - human-feedback --- # Dataset Card for Trisha This dataset has been created with [Argilla](https://github.com/argilla-io/argilla). As shown in
Hugging Face
2024-11-22 更新
4
0
DCAgent2/eval-DCAgent_exp_tas_max_tokens_1024_traces_DCAgent_dev_set_v2
对话评估
--- dataset_info: features: - name: conversations list: - name: content dtype: string - name: role dtype: string - name: agent dtype: string - name: model dtype
Hugging Face
2026-03-20 更新
6
0
autoevaluate/autoeval-staging-eval-project-6fbfec76-7855037
对话评估
模型评估
--- type: predictions tags: - autotrain - evaluation datasets: - samsum eval_info: task: summarization model: jpcorb20/pegasus-large-reddit_tifu-samsum-512 metrics: [] dataset_name: samsum d
Hugging Face
2022-06-27 更新
5
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广