← 返回专题广场
table-extraction
7 个项目 · ⭐ 32.9k1
16 天前
最近推送
2
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
Python
⭐ 10.5k
⑂ 785
AGPL-3.0
· 12 小时前推送
12 小时前
最近推送
3
6 小时前
最近推送
4
An on-premises, OCR-free unstructured data extraction, markdown conversion and benchmarking toolkit. (https://idp-leaderboard.org/)
Python
⭐ 2.1k
⑂ 155
Apache-2.0
· 2026-03-17推送
2026-03-17
最近推送
5
Pure Rust PDF library for AI/RAG: structure-aware chunking, no ML, no C deps.
Rust
⭐ 185
⑂ 26
MIT
· 4 小时前推送
4 小时前
最近推送
6
15 小时前
最近推送
7
2025-02-23
最近推送