youssefkhalil320/urdu_images_doc_tags_all_v11
收藏官方服务:
资源简介:
该数据集包含了文档相关的多种信息,包括文档ID、PDF路径、图像、图像预览、HTML和OTS格式标签、Markdown文本、语言类型及其识别置信度、难度评分、文本长度、是否含有缺失或不完整的边界框、页面尺寸以及渲染尺寸等。数据集分为训练集,并提供了相应的字节大小和示例数量。
The dataset includes various document-related information such as document ID, PDF path, images, image previews, HTML and OTS format tags, Markdown text, language type and its recognition confidence, difficulty score, text length, presence of missing or incomplete bounding boxes, page dimensions, and rendering dimensions. The dataset is split into a training set, with provided byte size and number of examples.
提供机构:
youssefkhalil320搜集汇总
数据集介绍

背景与挑战
背景概述
该数据集包含约6,741个乌尔都语文档页面图像及其对应的结构化标签(doctag),每个样本包含图像、文本位置信息和语言标签(乌尔都语)。适用于乌尔都语文档布局分析、OCR或文档理解等任务。
以上内容由遇见数据集搜集并总结生成



