P
PaLM-rlhf-pytorch
lucidrains/PaLM-rlhf-pytorch
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
★7.9kstars
Python
MIT
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?A artificial-intelligence/attention-mechanisms tool in the Models category, built with Python, open-source
Who made it?Maintained by lucidrains team, 7.9K⭐ on GitHub, #195 out of 3201 in Models
Why does it exist?With the growing demand for artificial-intelligence and attention-mechanisms in Models, lucidrains created PaLM-rlhf-pytorch as a streamlined solution.
What can it do?Key use cases: deep-learning, human-feedback, reinforcement-learning
How to install with AI?Use an AI coding assistant (Claude Code, Cursor, Copilot) to automatically set up pip dependencies and virtual env. Follow the README — the AI handles the rest.
🔗 github.com/lucidrains/PaLM-rlhf-pytorch
🔗 github.com/lucidrains/PaLM-rlhf-pytorch
Topics
artificial-intelligenceattention-mechanismsdeep-learninghuman-feedbackreinforcement-learningtransformers