waybarrios/vllm-mlx
High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.
⭐ 1.5k
⑂ 212
Python
Apache-2.0
· 3 小时前推送
1.5k
Watchers
0
贡献者
0
Commits
0
Releases
67
Open Issues
3 小时前
最近推送
原文
中文
暂无 README