Ejafa/ye-pop
收藏资源简介:
--- license: apache-2.0 language: - en tags: - art pretty_name: ye-pop size_categories: - 100K<n<1M --- # YE-POP (a derived dataset of Laion POP) YE-POP is a derived dataset from Laion-POP, meticulously curated and filtered to enhance the quality and utility of the original dataset. The dataset comprises 11 chunks, each containing 50,000 image URLs from Laion-POP. NSFW sorting has been used as a baseline, and human verification has been conducted to ensure the dataset's reliability. For the initial comparison, Chunk 1 has been curated with Gemini-Pro and released as part of a research work to the community. For access to other chunks generated by gemini-pro, interested parties are encouraged to contact us. The primary goal of YE-POP is to provide a dataset with improved art image descriptions while retaining the essence of Laion-POP for baseline comparisons in diffusion models and image captioning tasks. We anticipate that training multimodal models on this dataset will lead to enhanced generation capabilities. ## Dataset Details Each zip file contains predownloaded images, and the JSON file includes dictionaries of image features with the following fields: - `filename` - `url` - `cogvlm_caption` - `llava_caption` - `nsfw_prediction` - `alt_txt` - `alt_txt_similarity` - `width` - `height` - `original_width` - `original_height` - `exif` For more [detailed information](https://laion.ai/blog/laion-pop/#dataset-and-methodology) on the fields, refer to the JSON file. ## Dataset Card Authors [Yaroslav Ponomarenko]() [Ejafa Bassam]() ## Dataset Card Contact @[Peking University](https://cs.pku.edu.cn/English/Home.htm) ## Acknowledgments [Laion (Christoph Schuhmann, Peter Bevan)]() [Google Gemini-Pro](https://doi.org/10.48550/arXiv.2312.11805)
--- 许可证:Apache-2.0 语言:英语 标签:艺术 展示名称:YE-POP 规模类别:10万至100万条数据 --- # YE-POP(Laion POP衍生数据集) YE-POP 是 Laion-POP 的衍生数据集,经精心整理与筛选,旨在提升原始数据集的质量与应用价值。本数据集共包含11个数据块,每个数据块均包含来自Laion-POP的50000条图片URL。已以NSFW分类作为基准筛选流程,并通过人工核验保障数据集的可靠性。 首个用于对比测试的数据块(Chunk 1)已通过Gemini-Pro完成整理,并作为一项研究工作的组成部分向社区公开。若需获取由Gemini-Pro生成的其余数据块,请联系我们。YE-POP的核心目标是提供具备更优质艺术图像描述的数据集,同时保留Laion-POP的核心特性,以供扩散模型与图像字幕任务中的基准对比研究使用。我们预期,在该数据集上训练多模态模型将有效提升模型的生成能力。 ## 数据集详情 每个压缩包均包含预下载的图片,JSON文件则包含图像特征字典,其字段如下: - `filename`:文件名 - `url`:图片URL - `cogvlm_caption`:CogVLM图像描述 - `llava_caption`:LLaVA图像描述 - `nsfw_prediction`:NSFW预测结果 - `alt_txt`:替代文本 - `alt_txt_similarity`:替代文本相似度 - `width`:图片宽度 - `height`:图片高度 - `original_width`:原始图片宽度 - `original_height`:原始图片高度 - `exif`:EXIF信息 如需了解各字段的详细说明,请参阅JSON文件或访问[官方说明](https://laion.ai/blog/laion-pop/#dataset-and-methodology)。 ## 数据集卡片作者 [雅罗斯拉夫·波诺马连科(Yaroslav Ponomarenko)]() [埃贾法·巴萨姆(Ejafa Bassam)]() ## 数据集卡片联系方式 @[北京大学](https://cs.pku.edu.cn/English/Home.htm) ## 致谢 [Laion团队(克里斯托夫·舒曼、彼得·贝万)]() [Google Gemini-Pro](https://doi.org/10.48550/arXiv.2312.11805)
YE-POP 数据集概述
基本信息
- 许可证: Apache-2.0
- 语言: 英语
- 标签: 艺术
- 美观名称: ye-pop
- 大小分类: 100K<n<1M
数据集描述
YE-POP 是从 Laion-POP 派生的数据集,经过精心筛选和优化,以提高原始数据集的质量和实用性。该数据集包含11个部分,每个部分包含来自 Laion-POP 的50,000个图像URL。使用了NSFW分类作为基准,并通过人工验证确保数据集的可靠性。
数据集内容
每个zip文件包含预下载的图像,JSON文件包含图像特征的字典,具有以下字段:
filenameurlcogvlm_captionllava_captionnsfw_predictionalt_txtalt_txt_similaritywidthheightoriginal_widthoriginal_heightexif
数据集用途
YE-POP 的主要目标是提供一个改进的艺术图像描述数据集,同时保留 Laion-POP 的本质,用于扩散模型和图像字幕任务的基准比较。预计在此数据集上训练的多模态模型将提高生成能力。




