trlx
by CarperAI
Pythonpushed over 2 years ago
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
AI summary
RL Framework
A framework for distributed reinforcement learning of large language models with human feedback
- stars
- 4.5K
- forks
- 472
- watching
- 51
- awesome lists
- 2
Featured in 2 awesome lists
Each link jumps to the spot where the list mentions trlx.