Replies: 5 comments 1 reply
|
Can you please be more specific? Which tool to use? How to decide if at all to OCR? ... |
0 replies
|
Just convert image pdf to text bi-layer pdf with good layout |
0 replies
|
Try OCRmyPDF. If you want an integration within your PyMuPDF script use the integrated Tesseract access. There are example scripts in the utilities repo. |
0 replies
|
Tesseract is hard to train and is low in accuracy. I want to change to paddleocr, any method? |
1 reply
|
ok |
0 replies
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Uh oh!
There was an error while loading. Please reload this page.
What is the suggested approach for ocr pdf?
All reactions