vwxyzjn / lm-human-preference-details

RLHF implementation details of OAI's 2019 codebase
MIT License
152 stars 7 forks source link

Summarization TL;DR #27

Open vwxyzjn opened 1 year ago

vwxyzjn commented 1 year ago

Still WIP