Robert Sachunsky
|
c01609ff4e
|
allow even more empty imports for optional dependencies
|
2 weeks ago |
Robert Sachunsky
|
51e9bfd6d7
|
improve+extend dockerfile
|
2 weeks ago |
Robert Sachunsky
|
09248d4829
|
improve+extend makefile
|
2 weeks ago |
Robert Sachunsky
|
46618f4229
|
allow more empty imports for optional dependencies
|
2 weeks ago |
Robert Sachunsky
|
4be89910a2
|
CLI: fix arg vs kwarg from merge
|
2 weeks ago |
Robert Sachunsky
|
9d61acf173
|
simplify
|
2 weeks ago |
Robert Sachunsky
|
a1068ff2eb
|
OCR-D: move sbb-binarize to ocrd-tool.json, update to v3
|
2 weeks ago |
Robert Sachunsky
|
c794d4d29f
|
OCR-D: fix typo light_mode→light_version
|
2 weeks ago |
Robert Sachunsky
|
4338259ca1
|
OCR-D: ensure page image gets replaced in result as well if not the original file
|
2 weeks ago |
Robert Sachunsky
|
55969b0173
|
OCR-D: add docstring
|
2 weeks ago |
Robert Sachunsky
|
3916474b8b
|
OCR-D: require >=v3.1
|
2 weeks ago |
Robert Sachunsky
|
6d02e90570
|
OCR-D: restrict max_workers=1
|
2 weeks ago |
Robert Sachunsky
|
efd3fa6775
|
allow empty imports for optional dependencies
|
2 weeks ago |
Robert Sachunsky
|
238132e260
|
use 'image_filename' for pseudo-iteration outside 'dir_in' mode
|
2 weeks ago |
Robert Sachunsky
|
af4e2a4ffc
|
do not require 'dir_out' outside 'dir_in' mode
|
2 weeks ago |
Robert Sachunsky
|
ea136e3ddd
|
'overwrite' check: only in 'dir_in' mode
|
2 weeks ago |
Robert Sachunsky
|
1f4a17b60d
|
Merge remote-tracking branch 'origin/machine_based_reading_order_integration' into v3-api
|
2 weeks ago |
Robert Sachunsky
|
edf924c2cb
|
ocrd-tool: add dockerhub
|
2 weeks ago |
vahidrezanezhad
|
d3a4c06e7f
|
This commit enables the export of cropped text line images along with their corresponding texts from a Page-XML file. These exported text line images and texts can be utilized for training a text line-based OCR model.
|
4 weeks ago |
vahidrezanezhad
|
c8b8529951
|
For the CNN-RNN OCR model, long text lines are split into two segments
|
4 weeks ago |
vahidrezanezhad
|
aa72ca3006
|
Resolved an issue in the OCR-D framework where dir_out received a None value
|
1 month ago |
vahidrezanezhad
|
a4f1f35125
|
Resolving test failure
|
1 month ago |
kba
|
54040c1db4
|
Merge remote-tracking branch 'bertsky/machine_based_reading_order_integration_fixes' into machine_based_reading_order_integration
|
1 month ago |
vahidrezanezhad
|
7110bd971f
|
resolved an error for light version in the case that slope_deskew is smaller than slope_threshold
|
2 months ago |
vahidrezanezhad
|
25116a2c79
|
resolved 2 errors
|
2 months ago |
kba
|
869110f185
|
merge main
|
3 months ago |
vahidrezanezhad
|
33fda2f8be
|
changing cnn ocr model name
|
4 months ago |
Robert Sachunsky
|
335aa273a1
|
simplify, wrap extremely long lines
|
4 months ago |
Robert Sachunsky
|
cfc65128b1
|
reduce redundancy/indentation
|
4 months ago |
Robert Sachunsky
|
01376af905
|
do_order_of_regions_with_model: simplify
|
4 months ago |
vahidrezanezhad
|
92bfac4b41
|
Provide OCR as an option to process a directory of XML files, incorporating layout and text line coordinates.
|
4 months ago |
vahidrezanezhad
|
fbeef79d50
|
adding scatter_nd inference
|
4 months ago |
Robert Sachunsky
|
0ae28f7d3e
|
switch from stdlib to loky.ProcessPoolExecutor, ensure shutdown
|
4 months ago |
vahidrezanezhad
|
f93c6c288d
|
function of patch-wise inference with scatter_nd is added
|
4 months ago |
vahidrezanezhad
|
0e8c561618
|
debugging issues
|
4 months ago |
Robert Sachunsky
|
e9c0d716f6
|
CI: install optional dependencies, too
|
4 months ago |
Robert Sachunsky
|
dcaf796283
|
change polarity of orientation angle (PAGE schema required cw=positive)
|
4 months ago |
Robert Sachunsky
|
b4b0890294
|
add option to overwrite output xml, but skip by default if file exists
|
4 months ago |
Robert Sachunsky
|
b9ca7a6191
|
log num_cols-dependent resizing
|
4 months ago |
Robert Sachunsky
|
9270ea4550
|
annotate region angles in PAGE
|
4 months ago |
Robert Sachunsky
|
3b70b11ea6
|
avoid deskewing patches if binary-empty
|
4 months ago |
Robert Sachunsky
|
7e9ee90e6e
|
switch from (ad-hoc) mp.Pool to (attribute) concurrent.futures.ProcessPoolExecutor
|
4 months ago |
Robert Sachunsky
|
68456ea002
|
do_work_of_slopes_new*, do_back_rotation_and_get_cnt_back, do_work_of_contours_in_image: use mp.Pool, simplify
|
4 months ago |
Robert Sachunsky
|
25e967397d
|
exit early if no text regions found (to avoid segfault)
|
4 months ago |
Robert Sachunsky
|
21efea8711
|
no del on function argument
|
4 months ago |
Robert Sachunsky
|
5e0c1da711
|
simplify
|
4 months ago |
Robert Sachunsky
|
54cb15056b
|
do_image_rotation / return_deskew_slop: avoid code duplication, simplify via mp.Pool
|
4 months ago |
Robert Sachunsky
|
6fe02df973
|
do_image_rotation: fix f93fa12 (do return results)
|
4 months ago |
Robert Sachunsky
|
d68017037c
|
do_prediction: trigger GC to avoid CUDA OOM
|
4 months ago |
Robert Sachunsky
|
ad748d0039
|
do_prediction: avoid code duplication
|
4 months ago |