lucidrains

lucidrains/PaLM-rlhf-pytorch

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

⭐ 7.9k ⑂ 673 Python MIT · 26 天前推送
7.9k
Watchers
0
贡献者
0
Commits
0
Releases
20
Open Issues
26 天前
最近推送
原文 中文
暂无 README