zakariaf

zakariaf/RAG-Cache

High-performance LLM query cache with semantic search. Reduce API costs 80% and latency from 8.5s to 1ms using Redis + Qdrant vector DB. Multi-provider support (OpenAI, Anthropic).

⭐ 13 ⑂ 3 Python · 2025-12-02推送
13
Watchers
0
贡献者
0
Commits
0
Releases
18
Open Issues
2025-12-02
最近推送
原文 中文
暂无 README