Implementation of Memorizing Transformers (ICLR 2022), attention net augmented with indexing and retrieval of memories using approximate nearest neighbors, in Pytorch
In-memory vector store with efficient read and write performance for semantic caching and retrieval system. Redis for Semantic Caching.