登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
Fredithefish/hh-rlhf-RedPajama-Chat-Format
Fredithefish/hh-rlhf-RedPajama-Chat-Format
收藏
Hugging Face
2023-06-06 更新
2024-03-04 收录
基于人类反馈的强化学习
对话模型
数据链接:
https://hf-mirror.com/datasets/Fredithefish/hh-rlhf-RedPajama-Chat-Format
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
--- license: apache-2.0 ---
许可证:Apache-2.0
应用场景:
提供机构:
Fredithefish
原始信息汇总
数据集概述
授权信息
许可证
: Apache-2.0
相关数据集
bespokelabs/sft_data_nile_liyan_sharegpt
对话模型
工具调用对话
这是一个包含多个字段的数据集,其中包括提示和完成信息,每个字段都有指定的数据类型。数据集分为训练集,包含了196个示例。此外,还提供了数据集的配置信息。
Hugging Face
2025-05-30 更新
14
0
AIMH/SQPsychConv_llama3_no_questionnaire_finetune
心理咨询
对话模型
--- dataset_info: features: - name: file_id dtype: string - name: condition dtype: string - name: client_model dtype: string - name: therapist_model dtype: string - name: i
Hugging Face
2026-02-05 更新
9
0
thanhpn/iapp_wiki_qa_squad_oa
问答系统
对话模型
--- dataset_info: features: - name: INSTRUCTION dtype: string - name: RESPONSE dtype: string - name: SOURCE dtype: string splits: - name: train num_bytes: 1150840 num_e
Hugging Face
2023-07-04 更新
8
0
liyinghong/DFPO-Preft-taiyi
对话模型
强化学习对齐
--- license: mit dataset_info: features: - name: conversation_id dtype: int64 - name: category dtype: string - name: dataset dtype: string - name: language dtype:
Hugging Face
2024-10-23 更新
12
0
CodeDPO/rlhf_dataset_20250126_openrlhf_format
代码生成模型优化
基于人类反馈的强化学习
该数据集包含了问题、测试用例、推断和上下文消息等字段。它通过Qwen Coder 32B Instruct的过滤,用于测试用例和准确度的研究。数据集分为训练集,共有84924个示例,大小为1,136,916,895字节。
Hugging Face
2025-01-26 更新
9
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广