Awesome Lists

LLaVA-RLHF

by llava-rlhf

Pythonpushed almost 3 years ago

Aligning LMMs with Factually Augmented RLHF

AI summary

Reward alignment system

Aligns large multimodal models with factually enhanced reward functions to improve performance and mitigate hacking in reinforcement learning

stars
328
forks
24
watching
9

Add a GitHub project

Missing a project or an awesome list? Paste its GitHub URL and we fetch it right away.