Twitter-X-Grok-Edit-Image-Pair-instruction-Dataset
收藏资源简介:
Twitter-X-Grok-Edit-Image-Pair-instruction-Dataset 是一个精心策划的数据集,包含从 X(Twitter)上分享的 Grok 编辑或其他 AI 编辑中收集的图像编辑示例。数据集提供了对齐的编辑指令、源图像、参考图像(如提供)以及最终编辑后的输出。该数据集专为指令引导的图像编辑、多模态对齐和基于扩散的编辑任务而设计。数据集结构包括每个样本的 id(文件名/样本标识符)、control1_image(编辑前的基图像)、control2_image(用于指导编辑的参考图像,如适用)、target_image(编辑后的最终图像)和 instruction(描述变换的编辑指令)。数据集适用于基于指令的图像编辑模型、扩散编辑流程、视觉语言模型训练和多模态对齐研究。数据集中的图像经过手动整理,且大多数指令为手动编写。未来该数据集将继续扩展更新。
The Twitter-X-Grok-Edit-Image-Pair-instruction-Dataset is a carefully curated dataset comprising image editing examples collected from Grok edits or other AI-generated edits shared on X (Twitter). The dataset offers aligned editing instructions, source images, reference images (if provided), and final edited outputs. It is developed specifically for instruction-guided image editing, multimodal alignment, and diffusion-based editing tasks. Each sample in the dataset follows a structured format including: id (filename/sample identifier), control1_image (base image prior to editing), control2_image (reference image for guiding the editing process when applicable), target_image (final post-editing image), and instruction (editing directive describing the intended transformation). This dataset is applicable for training instruction-based image editing models, constructing diffusion editing pipelines, training vision-language models, and carrying out multimodal alignment research. All images within the dataset have been manually curated, and the majority of the editing instructions were manually written. This dataset will continue to be expanded and updated in the future.
Twitter-X-Grok-Edit-Image-Pair-instruction-Dataset 概述
数据集简介
Twitter-X-Grok-Edit-Image-Pair-instruction-Dataset 是一个从 X(Twitter)上分享的 Grok 编辑或其他 AI 编辑中收集并整理的图像编辑示例数据集。该数据集包含与源图像、参考图像(如提供)和最终编辑输出对齐的编辑指令。它专为指令引导的图像编辑、多模态对齐和基于扩散的编辑任务而设计,并将在未来更新中持续扩展。
数据集结构
数据集包含一个训练集(train),共有 616 个样本,总大小为 726,215,606 字节。
数据特征
每个样本包含以下字段:
- id:文件名/样本标识符(字符串类型)。
- control1_image:基础图像(编辑前)(图像类型)。
- control2_image:用于指导编辑的参考图像(在适用的样本中存在,应按原样使用)(图像类型)。
- target_image:编辑后的(最终)图像(图像类型)。
- instruction:描述变换的编辑指令(字符串类型)。
典型使用格式
(control1_image + instruction + 可用的 control2_image)→ target_image
数据文件
数据集以两个 .parquet 文件的形式分布在 data/ 目录中。包含 control2_image 的样本已在其相应条目中提供,应直接作为编辑条件的一部分使用。
预期用途
适用于:
- 基于指令的图像编辑模型。
- 扩散编辑流程。
- 视觉-语言模型训练。
- 多模态对齐研究。
重要说明
- 部分样本包含一个作为编辑条件组成部分的参考图像(
control2_image),当存在时应按原样使用。 - 大多数指令是手动编写的。
- 图像是从公开分享的 Grok 编辑中手动整理的。
- Grok 水印或任何其他视觉 AI 水印已从 target_image 中移除。
- 数据集将在未来版本中持续增长。




