Document Layout Analysis
-
Updated
Oct 30, 2024 - Python
Document Layout Analysis
Extract the MODS/ALTO metadata of a bunch of METS/ALTO files into pandas DataFrames for data analysis
Digitalized Collections of the Berlin State Library: ALTO-XML Processing Tools / batch NER + EL / BERT-pre-training
Add a description, image, and links to the qurator topic page so that developers can more easily learn about it.
To associate your repository with the qurator topic, visit your repo's landing page and select "manage topics."