H
hh-rlhf
anthropics/hh-rlhf
Human preference data for "Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback"
★1.9kstars
MIT
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?A security/security tool in the Security & Governance category, open-source
Who made it?Maintained by anthropics team, 1.9K⭐ on GitHub, #94 out of 413 in Security & Governance
Why does it exist?With the growing demand for security and security in Security & Governance, anthropics created hh-rlhf as a streamlined solution.
What can it do?Key use cases: security
How to install with AI?Use an AI coding assistant to follow the README and automatically handle the install and environment setup.
🔗 github.com/anthropics/hh-rlhf | 官网 https://arxiv.org/abs/2204.05862
🔗 github.com/anthropics/hh-rlhf | 官网 https://arxiv.org/abs/2204.05862