alignment-handbook
by huggingface
Robust recipes to align language models with human and AI preferences
AI summary
Alignment Recipes
Provides recipes and guidelines for training language models to align with human preferences and AI goals
- stars
- 4.8K
- forks
- 417
- watching
- 112
Similar projects
Found by comparing what the projects do, not just their names.
Datasets
A curated collection of high-quality datasets for training large language models.
Prompt papers
A curated list of papers on prompt-based tuning for pre-trained language models, providing insights and advancements in the field.
Model alignment
This project demonstrates the effectiveness of reinforcement learning from human feedback (RLHF) in improving small language models like GPT-2.
language model tuning
Training methods and tools for fine-tuning language models using human preferences
huggingface/trl10.3K
Transformer model trainer
A library designed to train transformer language models with reinforcement learning using various optimization techniques and fine-tuning methods.
Robotics toolkit
A platform providing pre-trained models, datasets, and tools for robotics with focus on imitation learning and reinforcement learning.
huggingface/peft16.7K
Parameter adaptation
An efficient method for fine-tuning large pre-trained models by adapting only a small fraction of their parameters
haotian-liu/llava20.7K
Visual Instruction System
A system that uses large language and vision models to generate and process visual instructions
Language model developers
Develops and maintains large language models with improved stability and performance
ML explanations
An explanation of key concepts and advancements in the field of Machine Learning
huggingface/transformers136.4K
Model repository
A collection of pre-trained machine learning models for various natural language and computer vision tasks, enabling developers to fine-tune and deploy these models on their own projects.
haifengl/smile6.1K
Machine learning library
A comprehensive machine learning framework that provides a wide range of algorithms and data structures for tasks such as classification, regression, clustering, and visualization.
GPT-4 data generator
This project generates instruction-following data using GPT-4 to fine-tune large language models for real-world tasks.
Text Generation Toolkit
A toolkit for deploying and serving Large Language Models (LLMs) for high-performance text generation
Machine learning automator
Automates machine learning workflows and optimizes model performance using large language models and efficient algorithms