登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
Appendix: Automatically Create Evolutionary Metaheuristic Algorithms Using Reinforcement Learning
Appendix: Automatically Create Evolutionary Metaheuristic Algorithms Using Reinforcement Learning
收藏
Figshare
2021-02-05 更新
2026-04-08 收录
自动化算法
强化学习
数据链接:
https://figshare.com/articles/dataset/Appendix_Automatically_Create_Evolutionary_Metaheuristic_Algorithms_Using_Reinforcement_Learning/13719853/1
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
- Data Descriptor AutoMH
应用场景:
创建时间:
2021-02-05
相关数据集
ScheduleNet: Learn to solve multi-agent scheduling problems with reinforcement learning
多智能体调度
强化学习
We propose ScheduleNet, a RL-based real-time scheduler, that can solve various types of multi-agent scheduling problems. We formulate these problems as a semi-MDP with episodic reward (makespan) and l
DataCite Commons
2025-11-20 更新
14
0
luca0621/multi-RLHF-processed-llama1B-dataset-with-1000-rewards
自然语言处理
强化学习
该数据集包含查询(query)、响应(response)和奖励(reward)三个特征,其中查询和响应为字符串类型,奖励为浮点数类型。数据集分为训练集和测试集,训练集包含8000个样本,测试集包含2000个样本。总下载大小为2276183字节,数据集总大小为7148705字节。
Hugging Face
2024-11-30 更新
7
0
FractalAIResearch/Fathom-V0.4-RL-Compression
学科文本生成
强化学习
Ramanujan-Ganit-R1-14B-RL是一个英文文本生成数据集,可能包含数学相关的文本内容,并在模型预训练后进行了后处理。具体的数据集内容和用途在README中未详细说明。
Hugging Face
2025-05-13 更新
6
0
CognitiveLab/hh-rlhf-formatted
强化学习
文本生成
--- dataset_info: features: - name: prompt dtype: string - name: chosen dtype: string - name: rejected dtype: string splits: - name: train num_bytes: 327229899 num_exam
Hugging Face
2024-01-18 更新
8
0
dataset_for_testing
机器人技术
强化学习
这是一个使用phospho开发工具生成的数据集,包含了一系列机器人和多个摄像头记录的片段。这个数据集可以直接用于通过模仿学习来训练策略,并且与LeRobot和RLDS兼容。
Hugging Face
2025-02-10 更新
6
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广