lucidrains/PaLM-rlhf-pytorch
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
⭐ 7.9k
⑂ 673
Python
MIT
· 26 天前推送
7.9k
Watchers
0
贡献者
0
Commits
0
Releases
20
Open Issues
26 天前
最近推送
原文
中文
暂无 README