Tamil Brahmi Stone Inscription
收藏资源简介:
The “Tamil Brahmi Stone Inscription Dataset (TBSI)** is the first open-source benchmark dataset dedicated to the processing and analysis of degraded Tamil stone inscriptions. Due to the lack of publicly available benchmark datasets for this specific domain, this repository bridges a critical gap in digital heritage and computational epigraphy. The dataset is compiled from the official archives of the **Tamil Nadu Department of Archaeology (TNDA)** and other publicly available epigraphic resources. Ancient Tamil stone inscriptions are invaluable historical records, but their automated analysis is severely hindered by natural degradation over centuries. This dataset is specifically curated to evaluate and train robust Computer Vision (CV) algorithms such as image enhancement, text segmentation, and Optical Character Recognition (OCR)—under real-world, highly degraded conditions.
「泰米尔婆罗米石刻数据集(Tamil Brahmi Stone Inscription Dataset, TBSI)」是首个专注于退化泰米尔石刻处理与分析的开源基准数据集。 鉴于该细分领域尚无公开可用的基准数据集,本数据集填补了数字遗产与计算金石学领域的关键空白。该数据集采集自泰米尔纳德邦考古部门(Tamil Nadu Department of Archaeology, TNDA)的官方档案,以及其他公开可获取的金石学资源。 古代泰米尔石刻是极具价值的历史遗存,但历经数百年的自然退化,其自动化分析工作受到严重制约。本数据集专为在真实且高度退化的实际场景中,评估与训练鲁棒的计算机视觉(Computer Vision, CV)算法(涵盖图像增强、文本分割与光学字符识别(Optical Character Recognition, OCR))而精心构建。




