OCR: Difference between revisions
No edit summary |
|||
Line 19: | Line 19: | ||
[[OCR/Results]] | [[OCR/Results]] | ||
[[Category:OCR]] | [[Category:OCR]] | ||
[[Category:Artificial Intelligence]] |
Revision as of 13:43, 19 February 2018
|
---|
Current state
- On-demand OCR extraction of ingredients using Tesseract 2 (production) and Tesseract 3 (.net) and Google Cloud Vision
- Uses the French dictionary for all languages
-- /home/off-fr/cgi# grep get_ocr * Ingredients.pm:use Image::OCR::Tesseract 'get_ocr'; Ingredients.pm: $text = decode utf8=>get_ocr($image,undef,'fra');
- Has a small custom dictionary for French ( /usr/share/tesseract-ocr/tessdata/fra.user-words)