官方服务:
资源简介:
Single character Urdu OCR data
单一字符乌尔都文光学字符识别数据集
应用场景:
创建时间:
2021-05-27
相关数据集
abdur75648/UTRSet-Real
UTRSet-Real数据集是一个专门为印刷体乌尔都语OCR研究设计的手动注释数据集。它包含超过11,000张印刷文本行图像,每张图像都经过精心注释。数据集的一个显著特点是其多样性,包括字体、文本大小、颜色、方向、光照条件、噪声、样式和背景的变化。这种多样性使其非常适合于训练和评估在现实世界中乌尔都语文本识别任务中表现出色的模型。此外,UTRSet-Synth数据集是一个高质量合成数据集,与UTR
Hugging Face2024-01-30 更新380
FIPU-OCR-CHAR: Font-Invariant Printed Urdu Character Dataset
The FIPU-OCR-CHAR dataset is a large-scale, font-invariant corpus of printed Urdu characters designed to support research in optical character recognition, font generalization, and script analysis. Th
Mendeley Data50
ULRs/Urdu-Newspaper-Benchmark
该数据集包含高分辨率和低分辨率的图像对以及相应的文本转录。数据集分为训练集,共有829个样本,适用于图像处理和文本分析等任务。
Hugging Face2025-08-01 更新50
abdur75648/UTRSet-Synth
--- title: UrduSet-Synth (UTRNet) emoji: 📖 colorFrom: red colorTo: green license: cc-by-nc-4.0 task_categories: - image-to-text language: - ur tags: - ocr - text recognition - urdu-ocr - utrnet prett
Hugging Face2024-01-30 更新140
RAVI: Synthetic Urdu Text Image Dataset for OCR
The RAVI dataset is a synthetic image dataset designed to support the development and training of Urdu OCR (Optical Character Recognition) models. It consists of 99,000 high-resolution images (256x256
NIAID Data Ecosystem50



