align-anything

Model aligner

Aligns large multimodal models with human intentions and values using various algorithms and fine-tuning methods.

Align Anything: Training All-modality Model with Feedback

GitHub

270 stars
9 watching
53 forks
Language: Python
last commit: almost 2 years ago
chameleondpolarge-language-modelsmultimodalrlhfvision-language-model

Related projects:

RepositoryDescriptionStars
pku-yuangroup/languagebindExtending pretraining models to handle multiple modalities by aligning language and video representations751
aidc-ai/ovisAn MLLM architecture designed to align visual and textual embeddings through structural alignment575
ethanyanjiali/minchatgptThis project demonstrates the effectiveness of reinforcement learning from human feedback (RLHF) in improving small language models like GPT-2.214
lancopku/iaisThis project proposes a novel method for calibrating attention distributions in multimodal models to improve contextualized representations of image-text pairs.30
x-plug/cvaluesEvaluates and aligns the values of Chinese large language models with safety and responsibility standards481
rlhf-v/rlhf-vAligns large language models' behavior through fine-grained correctional human feedback to improve trustworthiness and accuracy.245
pku-yuangroup/video-benchEvaluates and benchmarks large language models' video understanding capabilities121
pkunlp-icler/pca-evalAn open-source benchmark and evaluation tool for assessing multimodal large language models' performance in embodied decision-making tasks99
cmesher/inuktitutalignerdataScripts for aligning laboratory speech production data in Inuktitut3
pku-yuangroup/moe-llavaA large vision-language model using a mixture-of-experts architecture to improve performance on multi-modal learning tasks2,023
prosodylab/prosodylab-alignerTools for aligning laboratory speech production data to forced audio alignment using HTK and SoX.333
mshukor/evalign-iclEvaluating and improving large multimodal models through in-context learning21
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
jcgood/rosetta-panglossA Python library that uses machine learning and natural language processing to improve translation accuracy by aligning source and target languages0
opengvlab/multi-modality-arenaAn evaluation platform for comparing multi-modality models on visual question-answering tasks478