Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
Catalog snapshot
Repository health snapshot: 673 forks and 20 open issues at review time. Platform support is inferred conservatively from repository topics.
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM Fossfits includes PaLM-rlhf-pytorch as an actively maintained llm infra option: its public repository was updated on 2026-07-27, uses the MIT license, and had 7,863 stars when reviewed. Check its documentation and release notes before production adoption.
artificial intelligence · attention mechanisms · deep learning · human feedback · reinforcement learning