sidd21sharma/newDataSetForDemo
收藏资源简介:
--- dataset_info: features: - name: text dtype: string splits: - name: train num_bytes: 15409089 num_examples: 9846 - name: test num_bytes: 815811 num_examples: 518 download_size: 9461517 dataset_size: 16224900 --- # Guanaco: Lazy Llama 2 Formatting This is the excellent [`timdettmers/openassistant-guanaco`](https://huggingface.co/datasets/timdettmers/openassistant-guanaco) dataset, processed to match Llama 2's prompt format as described [in this article](https://huggingface.co/blog/llama2#how-to-prompt-llama-2). Useful if you don't want to reformat it by yourself (e.g., using a script). It was designed for [this article](https://mlabonne.github.io/blog/posts/Fine_Tune_Your_Own_Llama_2_Model_in_a_Colab_Notebook.html) about fine-tuning a Llama 2 model in a Google Colab.
数据集概述
数据集特征
- 名称: text
- 数据类型: string
数据集分割
- 训练集
- 样本数: 9846
- 存储大小: 15409089字节
- 测试集
- 样本数: 518
- 存储大小: 815811字节
数据集大小
- 下载大小: 9461517字节
- 总大小: 16224900字节



