Tesseract training for a new font

后端未结

关注

 3  1957

日久生厌 2021-01-31 10:17

I\'m still new to Tesseract OCR and after using it in my script noticed it had a relatively big error rate for the images I was trying to extract text from. I came across Tesser

3条回答

广开言路 (楼主)

2021-01-31 10:45

For anyone that is still going to read this, you can use this tool to get a traineddata file of whichever font you want. After that move the traineddata file in your tessdata folder. To use tesseract with the new font in Python or any other language (I think?) put lang = "Font"as second parameter in image_to_string function. It improves accuracy significantly but can still make mistakes ofcourse. Or you can just learn how to train tesseract for a new font manually with this guide: http://pretius.com/how-to-prepare-training-files-for-tesseract-ocr-and-improve-characters-recognition/.

0 讨论(0)

查看其它3个回答
发布评论:

提交评论
- 加载中...