You cannot select more than 25 topics
Topics must start with a letter or number, can include dashes ('-') and can be up to 35 characters long.
Gerber, Mike 4525dd2a9c | 5 years ago | |
---|---|---|
.screenshots | 5 years ago | |
qurator | 5 years ago | |
.travis.yml | 5 years ago | |
README.md | 5 years ago | |
pytest.ini | 5 years ago | |
requirements.txt | 5 years ago | |
setup.py | 5 years ago |
README.md
dinglehopper
dinglehopper is an OCR evaluation tool and reads ALTO, PAGE and text files.
Goals
- Useful
- As an UI tool
- For an automated evaluation
- As a library
- Unicode support
Usage
dinglehopper some-document.gt.page.xml some-document.ocr.alto.xml
This generates report.html
and report.json
.
As a OCR-D processor:
ocrd-dinglehopper -m mets.xml -I OCR-D-GT-PAGE,OCR-D-OCR-TESS -O OCR-D-OCR-TESS-EVAL
This generates HTML and JSON reports in the OCR-D-OCR-TESS-EVAL
filegroup.