OCR: Difference between revisions
No edit summary |
|||
Line 1: | Line 1: | ||
== Current state == | == Current state == | ||
* OCR extraction of | * On-demand OCR extraction of ingredients using Tesseract 2 (production) and Tesseract 3 (.net) | ||
* Uses the French dictionary for all languages | * Uses the French dictionary for all languages | ||
<pre> | <pre> |
Revision as of 12:21, 24 January 2016
Current state
- On-demand OCR extraction of ingredients using Tesseract 2 (production) and Tesseract 3 (.net)
- Uses the French dictionary for all languages
-- /home/off-fr/cgi# grep get_ocr * Ingredients.pm:use Image::OCR::Tesseract 'get_ocr'; Ingredients.pm: $text = decode utf8=>get_ocr($image,undef,'fra');
- Has a small custom dictionary for French ( /usr/share/tesseract-ocr/tessdata/fra.user-words)