← 返回专题广场
llms-benchmarking
3 个项目 · ⭐ 1.6k1
🔥 A list of tools, frameworks, and resources for building AI web agents
Python
⭐ 1.5k
⑂ 208
NOASSERTION
· 2026-07-10推送
2026-07-10
最近推送
2
Python SDK for experimenting, testing, evaluating & monitoring LLM-powered applications - Parea AI (YC S23)
Python
⭐ 82
⑂ 13
Apache-2.0
· 2025-02-13推送
2025-02-13
最近推送
3
[IEEE T-BIOM] FaceXBench: Evaluating Multimodal LLMs on Face Understanding
Python
⭐ 20
⑂ 1
MIT
· 2026-01-16推送
2026-01-16
最近推送