instructGOOSE
by xrsrke
Implementation of Reinforcement Learning from Human Feedback (RLHF)
AI summary
RLHF framework
A framework for training language models using human feedback and reinforcement learning
- stars
- 171
- forks
- 21
- watching
- 5
Similar projects
Found by comparing what the projects do, not just their names.
RL simulator
A framework for simulating and evaluating reinforcement learning from human feedback methods
RL framework
A high-performance implementation of reinforcement learning training pipelines using JAX and PyTorch-like functionality
RL framework
A framework for implementing complex reinforcement learning algorithms with flexibility and ease of implementation
Behavior alignment tool
Aligns large language models' behavior through fine-grained correctional human feedback to improve trustworthiness and accuracy.
RL framework
An open-source reinforcement learning framework for autonomous driving tasks using the Carla-Simulator environment and Ray/Rllib libraries.
Model alignment
This project demonstrates the effectiveness of reinforcement learning from human feedback (RLHF) in improving small language models like GPT-2.
RL framework
A collection of implementations of reinforcement learning algorithms in MATLAB
RL framework
A framework for parallel population-based reinforcement learning
LLM trainer
A flexible RL training framework designed for large language models
FL Framework
Provides a framework and theoretical foundation for Federated Reinforcement Learning with Byzantine Resilience in distributed systems
kaixhin/rainbow1.6K
RL framework
A Python implementation of a deep reinforcement learning algorithm combining multiple techniques for improved performance in Atari games
RL framework
A Python library implementing state-of-the-art deep reinforcement learning algorithms for Keras and OpenAI Gym environments.
CARLA RL framework
Provides a framework for using CARLA as a reinforcement learning environment
RL Framework
An RL framework for building and training reinforcement learning models in Python
RL toolkit
Provides a unified toolkit for constructing, computing, and optimizing intrinsic reward modules in reinforcement learning