Awesome Lists

minChatGPT

by ethanyanjiali

Pythonpushed about 3 years ago

A minimum example of aligning language models with RLHF similar to ChatGPT

AI summary

Model alignment

This project demonstrates the effectiveness of reinforcement learning from human feedback (RLHF) in improving small language models like GPT-2.

stars
214
forks
28
watching
5

Add a GitHub project

Missing a project or an awesome list? Paste its GitHub URL and we fetch it right away.