遇见数据集

Columbia-NLP/SFT-orca-math-word-problems-200k

收藏
Hugging Face2024-04-29 更新2024-07-22 收录
官方服务:

资源简介:

--- dataset_info: features: - name: messages list: - name: content dtype: string - name: role dtype: string splits: - name: train_sft num_bytes: 230323736 num_examples: 200035 download_size: 81723663 dataset_size: 230323736 configs: - config_name: default data_files: - split: train_sft path: data/train_sft-* ---

The dataset includes a feature named messages, which is a list containing two sub-features: content and role, both of which are of string type. The dataset is split into a partition named train_sft, containing 200035 samples with a total size of 230323736 bytes. The download size of the dataset is 81723663 bytes. The dataset configuration is named default, and the data file path is data/train_sft-*.

提供机构:
Columbia-NLP
二维码
社区交流群
二维码
科研交流群
商业服务