CatchTheTornado

CatchTheTornado/text-extract-api

Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported models. Anonymize documents. Remove PII. Convert any document or picture to structured JSON or Markdown

⭐ 3.2k ⑂ 279 Python MIT · 2025-12-09推送
3.2k
Watchers
0
贡献者
0
Commits
0
Releases
47
Open Issues
2025-12-09
最近推送
原文 中文
暂无 README