← 返回专题广场
load-balancing
4 个项目 · ⭐ 7.8k1
15 小时前
最近推送
2
8 天前
最近推送
3
The lowest-overhead LLM router. Production-ready, highly available, one OpenAI-compatible endpoint in front of 80 providers and your own vLLM/SGLang — 0.76 µs per request, no I/O on the request path, cache-affinity routing, RBAC, budgets and a 13-screen UI in the binary.
Rust
⭐ 106
⑂ 97
Apache-2.0
· 2 天前推送
2 天前
最近推送