verl

LLM trainer

A flexible RL training framework designed for large language models

veRL: Volcano Engine Reinforcement Learning for LLM

GitHub

427 stars
8 watching
29 forks
Language: Python
last commit: almost 2 years ago

Related projects:

RepositoryDescriptionStars
volcengine/vescaleA PyTorch-based framework for training large language models in parallel on multiple devices679
luchris429/purejaxrlA high-performance implementation of reinforcement learning training pipelines using JAX and PyTorch-like functionality755
toni-sm/skrlA modular reinforcement learning library with support for various environments and frameworks588
matthiasplappert/keras-rlA Python library implementing state-of-the-art deep reinforcement learning algorithms for Keras and OpenAI Gym environments.8
tristandeleu/pytorch-maml-rlReplication of Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks in PyTorch for reinforcement learning tasks830
vpgtrans/vpgtransTransfers visual prompt generators across large language models to reduce training costs and enable customization of multimodal LLMs270
rle-foundation/rlexploreProvides a unified toolkit for constructing, computing, and optimizing intrinsic reward modules in reinforcement learning373
astooke/rlpytA modular and unified framework for implementing common deep reinforcement learning algorithms in PyTorch2,236
kaixhin/rainbowA Python implementation of a deep reinforcement learning algorithm combining multiple techniques for improved performance in Atari games1,591
enlite-ai/mazeAn RL framework for building and training reinforcement learning models in Python266
millionintegrals/velA collection of modular deep learning components that can be easily configured and reused in various applications.276
kunqian2025/reinforcement-learningA collection of implementations of reinforcement learning algorithms in MATLAB61
zuoxingdong/lagomA modular toolkit for rapid prototyping of reinforcement learning algorithms373
mushroomrl/mushroom-rlA Python library for reinforcement learning algorithms and environments.824
iffix/machinAn open-source reinforcement learning library for PyTorch, providing a simple and clear implementation of various algorithms.402