相关数据集
ioi-leaderboard/ioi-eval-sglang_mistralai_Codestral-22B-v0.1-new-prompt
该数据集是一个包含编程问题及其相关信息的集合,其中包括问题ID、子任务、提示信息、生成内容、代码、编程语言、解决方案数量、唯一标识符以及模型参数等字段。数据集分为训练集,包含2050个示例,总大小为21183971字节。
Hugging Face2025-03-03 更新110
DCAgent2/swebench_verified_random_100_folders_daVinci_Dev_32B_20260424_211638
这是一个多轮对话数据集,包含对话内容(content)和角色(role),以及代理(agent)、模型(model)、模型提供商(model_provider)、日期(date)、任务(task)、回合(episode)、运行ID(run_id)、试验名称(trial_name)、结果(result)、验证器输出(verifier_output)等元数据。数据集用于训练或评估对话系统,共有298个
Hugging Face2026-04-25 更新100
Digital Forensics 2023 - DF2023
The actual Digital Forensics 2023 (DF2023) dataset is version 2: https://zenodo.org/record/7326540?token=eyJhbGciOiJIUzUxMiIsImV4cCI6MTY4OTU0ODM5OSwiaWF0IjoxNjY4NjA3NDM4fQ.eyJkYXRhIjp7InJlY2lkIjo3MzI2
Zenodo2022-11-02 更新60
nuprl/agnostics-codeforces-cots
--- dataset_info: features: - name: idx dtype: int64 - name: source_id dtype: string - name: prompt dtype: string - name: response dtype: string - name: problem_statement
Hugging Face2026-03-13 更新100
An Empirical Evaluation of GitHub Copilot’s Code Suggestions
Artifact accompanying MSR 2022 paper titled "An Empirical Evaluation of GitHub Copilot’s Code Suggestions" by Nhan Nguyen and Sarah Nadi
Figshare2022-03-29 更新60



