instructGOOSE

RLHF framework

A framework for training language models using human feedback and reinforcement learning

Implementation of Reinforcement Learning from Human Feedback (RLHF)

GitHub

171 stars
5 watching
21 forks
Language: Jupyter Notebook
last commit: over 3 years ago
chatgpthuman-feedbackinstructgptreinforcement-learningrlhf

Related projects:

RepositoryDescriptionStars
tatsu-lab/alpaca_farmA framework for simulating and evaluating reinforcement learning from human feedback methods786
luchris429/purejaxrlA high-performance implementation of reinforcement learning training pipelines using JAX and PyTorch-like functionality755
horizonrobotics/alfA framework for implementing complex reinforcement learning algorithms with flexibility and ease of implementation306
rlhf-v/rlhf-vAligns large language models' behavior through fine-grained correctional human feedback to improve trustworthiness and accuracy.245
layssi/carla_ray_rlibAn open-source reinforcement learning framework for autonomous driving tasks using the Carla-Simulator environment and Ray/Rllib libraries.35
ethanyanjiali/minchatgptThis project demonstrates the effectiveness of reinforcement learning from human feedback (RLHF) in improving small language models like GPT-2.214
kunqian2025/reinforcement-learningA collection of implementations of reinforcement learning algorithms in MATLAB61
sjtu-marl/malibA framework for parallel population-based reinforcement learning507
volcengine/verlA flexible RL training framework designed for large language models427
flint-xf-fan/byzantine-federated-rlProvides a framework and theoretical foundation for Federated Reinforcement Learning with Byzantine Resilience in distributed systems85
kaixhin/rainbowA Python implementation of a deep reinforcement learning algorithm combining multiple techniques for improved performance in Atari games1,591
matthiasplappert/keras-rlA Python library implementing state-of-the-art deep reinforcement learning algorithms for Keras and OpenAI Gym environments.8
gokulnc/setting-up-carla-reinforcement-learningProvides a framework for using CARLA as a reinforcement learning environment95
enlite-ai/mazeAn RL framework for building and training reinforcement learning models in Python266
rle-foundation/rlexploreProvides a unified toolkit for constructing, computing, and optimizing intrinsic reward modules in reinforcement learning373