trlx

RL Framework

A framework for distributed reinforcement learning of large language models with human feedback

A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)

GitHub

5k stars
51 watching
472 forks
Language: Python
last commit: over 2 years ago
Linked from 2 awesome lists

machine-learningpytorchreinforcement-learning

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
huggingface/trlA library designed to train transformer language models with reinforcement learning using various optimization techniques and fine-tuning methods.10,308
lucidrains/palm-rlhf-pytorchAn implementation of RLHF on top of the PaLM architecture to enable human feedback in reinforcement learning for large language models.7,729
tju-drl-lab/ai-optimizerA next-generation deep reinforcement learning toolkit with libraries for multiagent, self-supervised, offline, and transfer/reinforcement learning4,848
google-deepmind/trflProvides building blocks for Reinforcement Learning agents using TensorFlow3,136
paddlepaddle/parlA high-performance distributed training framework for Reinforcement Learning3,296
thu-ml/tianshouA high-performance reinforcement learning library with modular interfaces and user-friendly APIs for building deep learning agents.8,069
iffix/machinAn open-source reinforcement learning library for PyTorch, providing a simple and clear implementation of various algorithms.402
p-christ/deep-reinforcement-learning-algorithms-with-pytorchPyTorch implementations of popular deep reinforcement learning algorithms and environments.5,669
rle-foundation/rlexploreProvides a unified toolkit for constructing, computing, and optimizing intrinsic reward modules in reinforcement learning373
tristandeleu/pytorch-maml-rlReplication of Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks in PyTorch for reinforcement learning tasks830
luchris429/purejaxrlA high-performance implementation of reinforcement learning training pipelines using JAX and PyTorch-like functionality755
eleutherai/gpt-neoxProvides a framework for training large-scale language models on GPUs with advanced features and optimizations.6,997
tatsu-lab/alpaca_farmA framework for simulating and evaluating reinforcement learning from human feedback methods786
rlcode/reinforcement-learningA collection of clean and minimal examples for various reinforcement learning algorithms3,433
yandexdataschool/practical_rlAn educational resource teaching practical reinforcement learning skills in Python using popular deep learning frameworks.5,952