Sketch2Code-hf
收藏资源简介:
Sketch2Code数据集包含731个人类绘制的草图,与来自Design2Code数据集的484个真实网页配对。该数据集旨在为视觉语言模型(VLMs)在将基本草图转换为网页设计原型方面提供基准测试。数据集的特征包括ID、草图图像、源HTML和源截图图像。数据集分为一个名为'train'的分割,包含731个样本。README还提到,网页中的所有图像都被替换为一个蓝色的占位图像(rick.jpg)。
The Sketch2Code dataset comprises 731 human-drawn sketches paired with 484 real web pages sourced from the Design2Code dataset. This dataset aims to provide a benchmark for visual language models (VLMs) when converting basic hand-drawn sketches into web design prototypes. The dataset features include ID, sketch images, source HTML, and source screenshot images. It is split into a 'train' partition containing 731 samples. The README also mentions that all images within the web pages have been replaced with a blue placeholder image (rick.jpg).
Sketch2Code 数据集概述
数据集信息
-
特征:
- id: 字符串类型
- sketch: 图像类型
- source_html: 字符串类型
- source_screenshot: 图像类型
-
拆分:
- train:
- 样本数量: 731
- 数据大小: 178,040,827 字节
- train:
-
下载大小: 136,077,581 字节
-
数据集大小: 178,040,827 字节
配置
- config_name: default
- data_files:
- split: train
- path: data/train-*
- data_files:
数据集描述
- Sketch2Code 数据集包含 731 个人类绘制的草图,与 484 个来自 Design2Code 数据集 的真实网页配对。
- 该数据集用于基准测试视觉-语言模型 (VLMs) 将基本草图转换为网页设计原型的能力。
- 所有网页中的图像均被替换为蓝色占位图像 (rick.jpg)。




