RepoStreet
lucidrains

lucidrains/palm-rlhf-pytorch

steady

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

Stars
7,863
Forks
674
Open issues
17
24h
-1
-0.0%
7d
+1
+0.0%
Refresh
2h

Star history (7 days)

Last checked
53m ago
Last pushed
3d ago
Next check
just now