← 返回专题广场
m1
6 个项目 · ⭐ 42.4k1
11 天前
最近推送
2
2 小时前
最近推送
3
2026-06-20
最近推送
4
2025-06-27
最近推送
5
Run larger LLMs with longer contexts on Apple Silicon by using differentiated precision for KV cache quantization. KVSplit enables 8-bit keys & 4-bit values, reducing memory by 59% with <1% quality loss. Includes benchmarking, visualization, and one-command setup. Optimized for M1/M2/M3 Macs with Metal support.
Python
⭐ 362
⑂ 13
NOASSERTION
· 2025-05-21推送
2025-05-21
最近推送
6
2025-06-04
最近推送