遇见数据集

NOMOCRAT Maltese OCR data set

收藏
Zenodo2026-08-11 更新2026-08-13 收录
官方服务:

资源简介:

Data set produced by the NOMOCRAT project, a project with the aim to create a visual text extraction model from Maltese language PDFs using layout analysis, optical character recognition, and document reading order determination. The pages were taken from public Maltese language PDFs that can be found in the dokumenti.mt repository. Size is too small to be used for training but can be used for evaluation.

提供机构:
Zenodo
创建时间:
2026-08-11
二维码
社区交流群
二维码
科研交流群
商业服务