Tripitaka Koreana in Han (TKH) Dataset, Multiple Tripitaka in Han (MTH) Dataset
收藏资源简介:
这些数据集包含扫描的佛教经典图像,附有真实标签,包括每个框中的真实字符边界框和真实字符。TKH数据集包含1000张图像,约320,000个字符实例和23,000行,适合作为历史文档中字符检测和识别研究的基准数据集。MTH数据集包含约500张更复杂、更具代表性的图像,来自中国八种不同的佛教经典版本,更具挑战性,支持历史文档图像的研究。
These datasets comprise scanned images of Buddhist scriptures, accompanied by ground truth labels that include the true character bounding boxes and the actual characters within each box. The TKH dataset contains 1,000 images, approximately 320,000 character instances, and 23,000 lines, making it a suitable benchmark for research on character detection and recognition in historical documents. The MTH dataset includes around 500 more complex and representative images derived from eight different versions of Chinese Buddhist scriptures, presenting greater challenges and supporting research on historical document imagery.
数据集概述
数据集名称
- TKH数据集:Tripitaka Koreana in Han (TKH) Dataset
- MTH数据集:Multiple Tripitaka in Han (MTH) Dataset
数据集来源
- 由华南理工大学深度学习和视觉计算实验室发布。
数据集用途
- 用于历史文献中汉字检测与识别的研究。
数据集内容
- TKH数据集:包含1,000张图像,约320,000个字符实例和23,000行。图像布局相对规则,字符大多统一。
- MTH数据集:包含约500张图像,来自中国八个不同的佛经版本,情况更为复杂,包含绘图和多个文本区域,同一行中字符大小不一。
数据集特点
- 包含扫描的佛经图像及其对应的标注,包括字符边界框和每个框中的字符。
- 使用垂直投影方法和光束搜索算法进行字符分割和边界框调整。
数据集下载
使用限制
- 仅限于非商业研究目的使用。
实验结果
- 在TKH数据集上测试了多种检测方法,包括R-FCN、Faster R-CNN、SSD、YOLO、TextBoxes、DMP-Nets和FEN等,结果显示在Table 4中。
- 在MTH数据集上测试了模型的鲁棒性和泛化能力,结果显示在Table 5中。
联系方式
- 如有问题,请联系作者:eehlyang@mail.scut.edu.cn 和 eelwjin@scut.edu.cn。




