PaLM-rlhf-pytorch

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

7.9k
Stars
+-5
Gained
-0.1%
Growth
Python
Language

🎭 Best For

🏷️ Topics & Ecosystem

artificial-intelligence attention-mechanisms deep-learning human-feedback reinforcement-learning transformers

📊 Activity

Latest commit: 2026-07-27. Over the past 193 days, this repository gained -5 stars (+-0.1% growth). Activity data is based on daily RepoPi snapshots of the GitHub repository.