CohenQu/Polaris-AceReason-Math-0.0-0.125
收藏官方服务:
资源简介:
该数据集包含了问题、答案和平均奖励三个字段,适用于训练机器学习模型,特别是那些涉及问题回答和奖励机制的任务。训练集共有4370个示例,数据大小为1653558字节。
The dataset includes three fields: problem, answer, and mean_reward, suitable for training machine learning models, especially those involving question answering and reward mechanisms. The training set consists of 4370 examples with a total size of 1653558 bytes.
提供机构:
CohenQu


