← 返回专题广场
tesseract
10 个项目 · ⭐ 177.1k1
Tesseract Open Source OCR Engine (main repository)
C++
⭐ 76.1k
⑂ 10.8k
Apache-2.0
· 14 小时前推送
14 小时前
最近推送
2
Pure Javascript OCR for more than 100 Languages 📖🎉🖥
JavaScript
⭐ 38.7k
⑂ 2.4k
Apache-2.0
· 2026-05-17推送
2026-05-17
最近推送
3
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
Python
⭐ 34.5k
⑂ 2.4k
MPL-2.0
· 17 小时前推送
17 小时前
最近推送
4
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
Python
⭐ 10.5k
⑂ 785
AGPL-3.0
· 9 小时前推送
9 小时前
最近推送
5
3 小时前
最近推送
6
A wrapper to work with Tesseract OCR inside PHP.
PHP
⭐ 3.0k
⑂ 554
MIT
· 2026-01-28推送
2026-01-28
最近推送
7
Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs
Python
⭐ 3.0k
⑂ 214
NOASSERTION
· 19 天前推送
19 天前
最近推送
8
2025-12-03
最近推送
9
🔍 Better text detection by combining multiple OCR engines (EasyOCR, Tesseract, and Pororo) with 🧠 LLM.
Python
⭐ 638
⑂ 39
MIT
· 2025-06-11推送
2025-06-11
最近推送