sci-ml/trl
Train transformer language models with reinforcement learning (SFT, DPO, GRPO)
-
trl-1.9.2~amd64python_single_target_python3_12 python_single_target_python3_13 python_single_target_python3_14
View
Download
Browse License: Apache-2.0 Overlay: stuff -
trl-1.9.1~amd64python_single_target_python3_12 python_single_target_python3_13 python_single_target_python3_14
View
Download
Browse License: Apache-2.0 Overlay: stuff
USE Flags
python_single_target_python3_12
* This flag is undocumented *
python_single_target_python3_13
* This flag is undocumented *
python_single_target_python3_14
* This flag is undocumented *

