azrtydxb/Fastllm-proxy
The lowest-overhead LLM router. Production-ready, highly available, one OpenAI-compatible endpoint in front of 80 providers and your own vLLM/SGLang — 0.76 µs per request, no I/O on the request path, cache-affinity routing, RBAC, budgets and a 13-screen UI in the binary.
⭐ 106
⑂ 97
Rust
Apache-2.0
· 2 天前推送
106
Watchers
0
贡献者
0
Commits
0
Releases
2
Open Issues
2 天前
最近推送
原文
中文
暂无 README