Design2Code-HARD
收藏资源简介:
该数据集包含80个来自Github Pages的额外困难的网页,旨在挑战最先进的(SoTA)多模态大语言模型(LLMs)将视觉设计转换为代码实现的能力。每个示例都是一对源HTML和截图({id}.html和{id}.png)。所有网页中的图像都被替换为一个占位符图像(rick.jpg)。此外,数据集还有一个“简单”版本,可以在指定的链接中找到。更多信息可以参考项目页面和论文。
This dataset contains 80 extra-difficult webpages sourced from GitHub Pages, designed to challenge state-of-the-art (SoTA) multimodal large language models (LLMs) in their capability to translate visual designs into executable code. Each example includes a pair of source HTML file and corresponding screenshot, with filenames formatted as {id}.html and {id}.png respectively. All images on these webpages have been replaced with a placeholder image named rick.jpg. Additionally, a "simple" version of this dataset is available via the designated link. For more information, please refer to the project page and the associated paper.
Design2Code-HARD 数据集概述
数据集描述
- 数据来源: 80个来自Github Pages的额外困难网页。
- 数据类型: 每个示例包含一对源HTML文件和对应的截图({id}.html 和 {id}.png)。
- 数据用途: 用于挑战当前最先进的多模态大型语言模型(LLMs),测试其将视觉设计转换为代码实现的能力。
数据集特点
- 图像替换: 所有网页中的图像均被替换为占位符图像(rick.jpg)。
相关资源
- 简单版本: 参见 Design2Code 简单版本。
- 项目页面: 更多信息请访问 Design2Code 项目页面。
- 论文: 相关研究论文可在 arXiv 查阅。




