OpenAlignment

Owner: Anthropic

Access: Freely available on Hugging Face under the MIT license with no access restrictions.

Safety use case

Training reward models for RLHF; studying the helpfulness-harmlessness tradeoff in LLM alignment.

Benchmark relevance

Relevant to: HH-RLHF

Source

Anthropic HH-RLHF on Hugging Face

Status: Verified. Last verified: 2026-06-13.

Want to list this dataset on the marketplace?

Data Canon takes no transaction fees. Contact us to update listing details or add new datasets.

Contact about this dataset