OpenLLMAI / OpenRLHF

An Easy-to-use, Scalable and High-performance RLHF Framework (70B+ PPO Full Tuning & Iterative DPO & LoRA & Mixtral)
https://openrlhf.readthedocs.io/
Apache License 2.0
1.71k stars 160 forks source link

Easy to miss bug that results in min_new_tokens not working #330

Closed yannikkellerde closed 1 week ago

yannikkellerde commented 1 week ago

Removed the extra space