tomaarsen

tomaarsen/attention_sinks

Extend existing LLMs way beyond the original training length with constant memory usage, without retraining

⭐ 735 ⑂ 44 Python Apache-2.0 · 2024-04-11推送
735
Watchers
0
贡献者
0
Commits
0
Releases
21
Open Issues
2024-04-11
最近推送
原文 中文
暂无 README