Groudtruth Images for testing character segmentation algorithms from palm leaf images
收藏资源简介:
This dataset consists of manually segmented grantha characters from palmleaflets.Grantha characters was the script used in southern India to write sanskrit language.This dataset will serve as ground truth for testing segmentation algorithms for extraction of characters from palm leaf images.The palm leaflets are taken from the manuscript "Isadi Upanishad" obtained from Oriental Research Institute under University of Kerala,India.It consists of two input palm leaflets and the corresponding cropped characters of each line arranged in different folders.
本数据集包含来自棕榈叶手稿单页的经人工分割的格兰塔文(Grantha)字符。格兰塔文是印度南部用于书写梵语的书写体系。本数据集将作为测试从棕榈叶图像中提取字符的分割算法的基准真值(ground truth)。该数据集的棕榈叶手稿取材自印度喀拉拉邦大学东方研究所所藏的《伊萨奥义书(Isadi Upanishad)》手稿。数据集包含两份原始棕榈叶手稿样本,以及按不同文件夹分类存放的对应每一行的裁剪字符图像。




