# Dataset Card for Chinese-dpo-pairs Well-curated 10K reference pairs in Chinese. Data are created by GPT-3.5 translation from multiple sources, including: - flan_v2, sharegpt, ultrachat, evol_inst
Dataset of [MegaStyle](https://jeoyal.github.io/MegaStyle/). MegaStyle-1.4M is a large-scale style dataset built through a scalable pipeline that leverages consistent text-to-image style mapping of Q
# Dataset Information A Chain of Thought (CoT) version of the TAT-QA arithmetic dataset (hosted at https://huggingface.co/datasets/nvidia/ChatQA-Training-Data). The dataset was synthetically generat