lucidrains/PaLM-rlhf-pytorch

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

Open original source ↗

What this repository does

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

Repository facts

Owner
lucidrains
Primary language
Python
Stars
7,862
Forks
673
Open issues + pull requests
20
License
MIT
Archived
No
Default branch
main
Created
2022-12-09T17:53:46.000Z
Last push
2026-07-27T23:39:41.000Z

Topics and intended use

Owner-supplied topics: artificial-intelligence, attention-mechanisms, deep-learning, human-feedback, reinforcement-learning, transformers

Review the README for scope, installation, examples and limitations. We do not run repository code or certify it.

Evaluate before installing

Review licensing and dependencies, inspect recent commits and unresolved issues, and test in an isolated environment before production use. Stars and forks alone cannot answer these questions.

README and project files

Issues and maintenance discussion

Releases and changelog

Source and freshness

Source: GitHub. Metadata observed 2026-09-15T06:38:52.964Z. Daily imports are snapshots, not real-time monitoring.

Popularity and source listings do not establish security, suitability, licensing rights or benchmark performance.

Open original source

Related repositories

public-apis/public-apis

EbookFoundation/free-programming-books

donnemartin/system-design-primer

vinta/awesome-python

practical-tutorials/project-based-learning

NousResearch/hermes-agent