遇见数据集

micahr234/play_cartpole

收藏
Hugging Face2026-05-06 更新2026-05-31 收录
官方服务:

资源简介:

该数据集是一个强化学习数据集,包含环境信息、动作、奖励、观察等特征,用于训练和评估强化学习模型。训练集包含1000万个样本,评估集包含20万个样本。

This dataset is a reinforcement learning dataset containing features such as environment information, actions, rewards, and observations, used for training and evaluating reinforcement learning models. The training set consists of 10 million samples, and the evaluation set consists of 200,000 samples.

提供机构:
micahr234
二维码
社区交流群
二维码
科研交流群
商业服务