← 返回专题广场
llm-eval
8 个项目 · ⭐ 48.5k1
4 小时前
最近推送
3
🐢 Open-Source Evaluation & Testing library for LLM Agents
Python
⭐ 5.8k
⑂ 522
Apache-2.0
· 2 天前推送
2 天前
最近推送
4
Evaluation and Tracking for LLM Experiments and AI Agents
Python
⭐ 3.5k
⑂ 329
MIT
· 1 天前推送
1 天前
最近推送
5
UpTrain is an open-source unified platform to evaluate and improve Generative AI applications. We provide grades for 20+ preconfigured checks (covering language, code, embedding use-cases), perform root cause analysis on failure cases and give insights on how to resolve them.
Python
⭐ 2.4k
⑂ 204
Apache-2.0
· 2024-08-18推送
2024-08-18
最近推送
6
A desktop MCP client designed as a tool unitary utility integration, accelerating AI adoption through the Model Context Protocol (MCP) and enabling cross-vendor LLM API orchestration.
TypeScript
⭐ 1.2k
⑂ 108
Apache-2.0
· 2026-05-14推送
2026-05-14
最近推送
7
Python SDK for experimenting, testing, evaluating & monitoring LLM-powered applications - Parea AI (YC S23)
Python
⭐ 82
⑂ 13
Apache-2.0
· 2025-02-13推送
2025-02-13
最近推送