← 返回专题广场
document-analysis
14 个项目 · ⭐ 89.6k1
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
Python
⭐ 78.2k
⑂ 6.6k
NOASSERTION
· 3 天前推送
3 天前
最近推送
2
A system for agentic LLM-powered data processing and ETL
Python
⭐ 4.0k
⑂ 426
MIT
· 12 天前推送
12 天前
最近推送
3
Read and extract text and other content from PDFs in C# (port of PDFBox)
C#
⭐ 2.5k
⑂ 329
Apache-2.0
· 11 小时前推送
11 小时前
最近推送
4
An on-premises, OCR-free unstructured data extraction, markdown conversion and benchmarking toolkit. (https://idp-leaderboard.org/)
Python
⭐ 2.1k
⑂ 155
Apache-2.0
· 2026-03-17推送
2026-03-17
最近推送
5
2026-03-17
最近推送
6
A package for parsing PDFs and analyzing their content using LLMs.
Python
⭐ 267
⑂ 11
MIT
· 2024-08-06推送
2024-08-06
最近推送
7
5 天前
最近推送
8
3 天前
最近推送
9
2024-07-05
最近推送
10
2026-04-06
最近推送
11
Open-source Legal AI workspace for evidence-grounded legal drafting, matter analysis and verifiable answers.
Python
⭐ 50
⑂ 14
Apache-2.0
· 2026-05-17推送
2026-05-17
最近推送
12
6 天前
最近推送
13
2025-11-23
最近推送