zakariaf/RAG-Cache
High-performance LLM query cache with semantic search. Reduce API costs 80% and latency from 8.5s to 1ms using Redis + Qdrant vector DB. Multi-provider support (OpenAI, Anthropic).
⭐ 13
⑂ 3
Python
· 2025-12-02推送
13
Watchers
0
贡献者
0
Commits
0
Releases
18
Open Issues
2025-12-02
最近推送
原文
中文
暂无 README