lucidrains/palm-rlhf-pytorch
steadyImplementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
Python
View on GitHub
Stars
7,863
Forks
674
Open issues
17
24h
-1
-0.0%
7d
+1
+0.0%
Refresh
2h
Star history (7 days)
Last checked
53m ago
Last pushed
3d ago
Next check
just now