PaLM-rlhf-pytorch
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
🎭 Best For
🏷️ Topics & Ecosystem
artificial-intelligence
attention-mechanisms
deep-learning
human-feedback
reinforcement-learning
transformers
📊 Activity
Latest commit: 2026-07-27. Over the past 193 days, this repository gained -5 stars (+-0.1% growth). Activity data is based on daily RepoPi snapshots of the GitHub repository.