← 返回专题广场
inference-server
14 个项目 · ⭐ 39.1k1
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Python
⭐ 20.3k
⑂ 1.7k
Apache-2.0
· 1 天前推送
1 天前
最近推送
2
2 天前
最近推送
3
12 小时前
最近推送
4
Open-source inference server and production cluster for all the models your agent needs.
Python
⭐ 2.8k
⑂ 272
Apache-2.0
· 1 天前推送
1 天前
最近推送
5
Turn any computer or edge device into a command center for your computer vision projects.
Python
⭐ 2.4k
⑂ 308
NOASSERTION
· 13 小时前推送
13 小时前
最近推送
6
2 小时前
最近推送
8
TurboOCR, >200 img/s OmnidocBench. TensorRT FP16, PP-OCRv6, HTTP + gRPC
C++
⭐ 1.0k
⑂ 98
MIT
· 2 天前推送
2 天前
最近推送
9
Python + Inference - Model Deployment library in Python. Simplest model inference server ever.
Python
⭐ 543
⑂ 83
Apache-2.0
· 2023-02-15推送
2023-02-15
最近推送
10
PMetal: high-performance Apple Silicon framework for local LLM inference, LoRA/QLoRA fine-tuning, serving, quantization, and MLX/Metal acceleration.
Rust
⭐ 309
⑂ 24
NOASSERTION
· 2026-06-05推送
2026-06-05
最近推送
11
[⛔️ DEPRECATED] Friendli: the fastest serving engine for generative AI
Python
⭐ 50
⑂ 7
Apache-2.0
· 2025-06-25推送
2025-06-25
最近推送
12
1 天前
最近推送
13
2026-06-23
最近推送
14
2026-04-27
最近推送