← all repositories

tesseract-ocr/tessdata

Repository hosting trained LSTM-based OCR models for Tesseract text recognition.

tessdata
Not currently ranked — collecting fresh signals.
star history

Provides pre-trained language data files for the Tesseract OCR engine. Includes integerized LSTM neural network models that balance speed and accuracy for text extraction from images. Supports the modern neural net engine (–oem 1) as well as legacy models for backward compatibility.

Frequently asked

What is tesseract-ocr/tessdata?
Repository hosting trained LSTM-based OCR models for Tesseract text recognition.
Is tessdata open source?
Yes — tesseract-ocr/tessdata is open source, released under the Apache-2.0 license.
How popular is tessdata?
tesseract-ocr/tessdata has 7.6k stars on GitHub.
Where can I find tessdata?
tesseract-ocr/tessdata is on GitHub at https://github.com/tesseract-ocr/tessdata.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.