datasets · openai
Human preference comparisons on news article summaries for RLHF reward-model training.
| Weight file SHA-256 | 6409923afa098c27161af8588dd17dc830bd711d3eb9598bb776815ce9e9872a |
|---|
| Publisher | openai |
|---|---|
| Origin | United States |
| Licence | MIT |
| Modality | Text |
| Pulls recorded here | 28,000 |
Registry record as of 2026-06-06. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.