ajibawa-2023/Code-290k-ShareGPT
收藏资源简介:
--- license: apache-2.0 task_categories: - conversational - text-generation language: - en tags: - code size_categories: - 100K<n<1M --- **Code-290k-ShareGPT** This dataset is in Vicuna/ShareGPT format. There are around 290000 set of conversations. Each set having 2 conversations. Along with Python, Java, JavaScript, GO, C++, Rust, Ruby, Sql, MySql, R, Julia, Haskell, etc. code with detailed explanation are provided. This datset is built upon using my existing Datasets [Python-Code-23k-ShareGPT](https://huggingface.co/datasets/ajibawa-2023/Python-Code-23k-ShareGPT) and [Code-74k-ShareGPT](https://huggingface.co/datasets/ajibawa-2023/Code-74k-ShareGPT) My Models [Python-Code-13B](https://huggingface.co/ajibawa-2023/Python-Code-13B) and [Python-Code-33B](https://huggingface.co/ajibawa-2023/Python-Code-33B) are trained on [Python-Code-23k-ShareGPT](https://huggingface.co/datasets/ajibawa-2023/Python-Code-23k-ShareGPT). My Models [Code-13B](https://huggingface.co/ajibawa-2023/Code-13B) and [Code-33B](https://huggingface.co/ajibawa-2023/Code-33B) are trained on [Code-74k-ShareGPT](https://huggingface.co/datasets/ajibawa-2023/Code-74k-ShareGPT). I am building few models using **Code-290k-ShareGPT** dataset.
数据集概述
数据集名称
- Code-290k-ShareGPT
数据集内容
- 该数据集名称为Code-290k-ShareGPT,具体内容未在README文件中详细描述。



