This collection contains the benchmark data used for benchmarking text extraction tools. The data conatins : - List of documents - Ground truth data for each document - Additional
This dataset is developed for the 4DHydro project (https://4dhydro.eu/) working package 2. The 4DHydro working package 2 is dedicated to provided a set of land-surface and hydrological model community