Anarchiviste/GenHisDoc_dataset
Preview • Updated • 89 • 1
the GenHisDoc dataset is on Huggin Face
GenHisDoc is a generalistic datasets for historical documents layout recognition and detection. GenHisDoc use a combination of several previously published datasets which have been adapted and re-annotated to work together and our own annotated data. All embedded datasets are licensed under Creative Commons or other licenses that permit reuse.
This repository contains the different models and training metrics on different versions of GenHisDoc. Every directory is a model with the weights, the test run and an info.yaml with information about the parameters used.