tomaarsen/attention_sinks
Extend existing LLMs way beyond the original training length with constant memory usage, without retraining
⭐ 735
⑂ 44
Python
Apache-2.0
· 2024-04-11推送
735
Watchers
0
贡献者
0
Commits
0
Releases
21
Open Issues
2024-04-11
最近推送
原文
中文
暂无 README