Tesseract Open Source OCR Engine (main repository)
3 things to get right when pulling text out of scanned Indian paperwork — the language pack, the scan quality, and when a free tool beats a paid one. Apache-2.0, no per-page fee.
Stars and forks are a snapshot taken when this directory was last refreshed, not a live count — the figure here matches the one in our articles and videos. Last refreshed 5 Sept 2026.