ArtmeScienceLab/WordCon-WC-dataset
收藏资源简介:
WordCon-WC数据集是一个用于训练和评估文本渲染与布局控制模型的数据集。它包含五个子文件夹(2、3、4、5、6),每个文件夹对应一种不同的字体类型,用于渲染图像中的文本。数据集还包括区域文本掩码(存储在rearranged_masks.zip文件中)以及元数据和注释文件(captions.json)。captions.json文件提供了完整的文本条件输入(prompt)、掩码单词列表(gt_content_mask)、掩码路径(gt_mask_path)、目标控制单词(gt_target)和对应掩码索引(gt_target_mask),支持模型在文本样式和布局方面的精细控制。
The WordCon-WC dataset is designed for training and evaluating text rendering and layout control models. It includes five sub-folders (2, 3, 4, 5, 6), each corresponding to a different font type used for rendering text in images. The dataset also contains regional text masks (stored in rearranged_masks.zip) and a metadata and annotations file (captions.json). The captions.json file provides complete text condition inputs (prompt), lists of masked words (gt_content_mask), mask paths (gt_mask_path), target control words (gt_target), and corresponding mask indices (gt_target_mask), enabling fine-grained control over text styling and layout in models.





