Columbia-NLP/SFT-orca-math-word-problems-200k
收藏官方服务:
资源简介:
--- dataset_info: features: - name: messages list: - name: content dtype: string - name: role dtype: string splits: - name: train_sft num_bytes: 230323736 num_examples: 200035 download_size: 81723663 dataset_size: 230323736 configs: - config_name: default data_files: - split: train_sft path: data/train_sft-* ---
The dataset includes a feature named messages, which is a list containing two sub-features: content and role, both of which are of string type. The dataset is split into a partition named train_sft, containing 200035 samples with a total size of 230323736 bytes. The download size of the dataset is 81723663 bytes. The dataset configuration is named default, and the data file path is data/train_sft-*.
提供机构:
Columbia-NLP


