The Rodrigo corpus
收藏NIAID Data Ecosystem2026-03-11 收录
数据链接:
官方服务:
资源简介:
The Rodrigo corpus was obtained from the digitisation of the book “Historia de España del arçobispo Don Rodrigo”, written in ancient Spanish in 1545. It is a single writer book where most pages consist of a single block of well-separated lines of calligraphical text. This dataset is free available for research purposes. It contains 15,010 images of text lines with their paleographic transcription. It is divided into three partitions: 9000 text lines for training, 1000 for validation and 5010 for testing.
创建时间:
2020-01-24




