← 返回专题广场
distributed-inference
5 个项目 · ⭐ 11.0k1
A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.
Python
⭐ 5.5k
⑂ 624
Apache-2.0
· 1 天前推送
1 天前
最近推送
2
Achieve state of the art inference performance with modern accelerators on Kubernetes
Shell
⭐ 4.1k
⑂ 694
Apache-2.0
· 10 小时前推送
10 小时前
最近推送
3
SGLang-Omni empowers high-performance serving for TTS, ASR, speech and omni models.
Python
⭐ 913
⑂ 368
Apache-2.0
· 4 小时前推送
4 小时前
最近推送
4
1 天前
最近推送
5
An imperative command-line-interface for AI workload orchestration
Python
⭐ 25
⑂ 3
Apache-2.0
· 2 天前推送
2 天前
最近推送