Akseltinfat/Tifinagh-OCR-39K
收藏资源简介:
该数据集是一个包含39,101个合成图像的全面集合,专为训练和评估Tifinagh脚本的OCR和视觉语言模型而设计。它包含了多种字体、背景颜色和文本样式的矩形格式图像。数据集中的每个样本都包含图像文件路径、真实的Tifinagh转录文本、背景颜色、文本颜色和唯一的图像标识符。该数据集适用于Tifinagh OCR训练、Amazigh语言文档理解、视觉语言模型验证以及跨不同排版样式的鲁棒性测试。
This dataset is a comprehensive collection of 39,101 synthetic images designed for training and evaluating OCR and vision-language models for the Tifinagh script. It features a wide variety of fonts, background colors, and text styles in a rectangular format. Each sample in the dataset contains the image file path, ground-truth Tifinagh transcription, background color, text color, and a unique image identifier. The dataset is suitable for Tifinagh OCR training, document understanding for Amazigh languages, vision-language model validation, and robustness testing across different typographic styles.




