hh-rlhf
Human preference data for "Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback"
파일 탐색기
최종 버전 다운로드 (.zip)- test.jsonl.gz
- train.jsonl.gz
- test.jsonl.gz
- train.jsonl.gz
- test.jsonl.gz
- train.jsonl.gz
- test.jsonl.gz
- train.jsonl.gz
- red_team_attempts.jsonl.gz
- .gitattributes
- LICENSE
- README.md
// repository documentation
Was this content helpful?
(0 ratings)
