Mercurial > hgrepos > Python2 > PyMuPDF
view mupdf-source/thirdparty/tesseract.txt @ 42:4621bd954a09
FIX: Need packaging at runtime to because of the parsing of mupdf_version into a tuple.
In the merge it was removed erroneously.
| author | Franz Glasner <fzglas.hg@dom66.de> |
|---|---|
| date | Sat, 11 Oct 2025 17:14:12 +0200 |
| parents | b50eed0cc0ef |
| children |
line wrap: on
line source
If you want to build with Tesseract functionality, you need to run make with a "tesseract=yes" argument. You will also need a suitable set of traineddata for the languages you wish to run. Only the LSTM engine (the latest and most accurate engine) is built into Tesseract, so the traineddata contained within the repository itself is no good. Suitable data can be retrieved from either: https://github.com/tesseract-ocr/tessdata_best or https://github.com/tesseract-ocr/tessdata_fast e.g. wget https://github.com/tesseract-ocr/tessdata_fast/raw/master/eng.traineddata
