遇见数据集

AgentCL

收藏
魔搭社区2026-06-03 更新2026-07-19 收录
官方服务:

资源简介:

# AgentCL: Evaluation Framework for Continual Learning in Agents ## Datasets - `agentboard_babyai`, `agentboard_scienceworld`, and `mmlu_pro` are included as ready-to-use subsets. They are direct subsets from existing public datasets, with no modifications to the original data. - Our constructed `codeeval-pro` and `browsecomp_plus` streams are provided. The metadata in source datasets only indicates the authorship for the source datasets, and does not imply the authorship of this repository. The source cards remain the authority for dataset terms and licenses. ### CodeEval-Pro We build CodeEval-Pro files from the following sources: - Paired-task inputs: ```text raw_public/humaneval_pro.json https://huggingface.co/datasets/CodeEval-Pro/humaneval-pro raw_public/mbpp_pro.json https://huggingface.co/datasets/CodeEval-Pro/mbpp-pro raw_public/bigcodebench_lite_pro.json https://huggingface.co/datasets/CodeEval-Pro/bigcodebench-lite-pro ``` - Raw-task test-code inputs: ```text raw_test_code/humaneval-raw-test-json https://huggingface.co/datasets/openai/openai_humaneval raw_test_code/mbpp-raw-test-json https://huggingface.co/datasets/evalplus/mbppplus raw_test_code/bigcodebench-raw-test-json https://huggingface.co/datasets/bigcode/bigcodebench ```

提供机构:
maas
创建时间:
2026-06-02
二维码
社区交流群
二维码
科研交流群
商业服务