alpaca_farm
by tatsu-lab
Pythonpushed about 2 years ago
A simulation framework for RLHF and alternatives. Develop your RLHF method without collecting human data.
AI summary
RL simulator
A framework for simulating and evaluating reinforcement learning from human feedback methods
- stars
- 786
- forks
- 59
- watching
- 9
- awesome list
- 1
Featured in 1 awesome list
Each link jumps to the spot where the list mentions alpaca_farm.