遇见数据集

OrionLLM/OpenAgentInstruct

收藏
Hugging Face2026-04-19 更新2026-05-10 收录
官方服务:

资源简介:

--- dataset_info: features: - name: conversations list: - name: content dtype: string - name: role dtype: string splits: - name: train num_bytes: 430031471 num_examples: 15209 download_size: 430095010 dataset_size: 430031471 configs: - config_name: default data_files: - split: train path: data/train-* license: apache-2.0 size_categories: - 10K<n<100K --- # Welcome to OpenAgentInstruct **OpenAgentInstruct** is a chat-style **Supervised Fine-Tuning (SFT)** dataset built for **terminal / command-line response**. - **Rows:** 15,209 (train split only) - **Size:** ~430 MB - **Language:** English (primarily)

## 数据集信息 --- 特征: - 名称:conversations,类型为列表,包含两个子字段: - content:数据类型为字符串(string) - role:数据类型为字符串(string) 数据拆分: - 名称:train,字节数:430031471,样本数:15209 下载大小:430095010 数据集存储大小:430031471 配置项: - 配置名称:default,数据文件: - 拆分:train,路径:data/train-* 许可证:Apache-2.0 大小分类: - 10K<n<100K --- # 欢迎使用 OpenAgentInstruct **OpenAgentInstruct** 是一款专为终端/命令行响应场景打造的对话式**监督微调(Supervised Fine-Tuning, SFT)**数据集。 - **样本条数**:仅训练拆分即包含15209条样本 - **数据集体量**:约430 MB - **主要语言**:以英语为主

提供机构:
OrionLLM
二维码
社区交流群
二维码
科研交流群
商业服务