REVO-LION

VLIT model toolkit

A comprehensive dataset and evaluation framework for Vision-Language Instruction Tuning models

REVO-LION: Evaluating and Refining Vision-Language Instruction Tuning Datasets

GitHub

11 stars
1 watching
0 forks
last commit: almost 3 years ago

Related projects:

RepositoryDescriptionStars
jiutian-vl/jiutian-lionThis project integrates visual knowledge into large language models to improve their capabilities and reduce hallucinations.124
ys-zong/vlguardImproves safety and helpfulness of large language models by fine-tuning them using safety-critical tasks47
jiasenlu/vilbert_betaA pre-trained model and toolset for performing vision-and-language tasks using a specific neural network architecture.473
zhuiyitechnology/pretrained-modelsA collection of pre-trained language models for natural language processing tasks989
vhellendoorn/code-lmsA guide to using pre-trained large language models in source code analysis and generation1,789
aidc-ai/parrotA method and toolkit for fine-tuning large language models to perform visual instruction tasks in multiple languages.34
nvlabs/prismerA deep learning framework for training multi-modal models with vision and language capabilities.1,299
vlf-silkie/vlfeedbackAn annotated preference dataset and training framework for improving large vision language models.88
byungkwanlee/collavoDevelops a PyTorch implementation of an enhanced vision language model93
ymcui/macbertImproves pre-trained Chinese language models by incorporating a correction task to alleviate inconsistency issues with downstream tasks646
yg-smile/rl_vvc_datasetA collection of benchmarks and implementations for testing reinforcement learning-based Volt-VAR control algorithms20
flagai-open/aquila2Provides pre-trained language models and tools for fine-tuning and evaluation439
yiren-jian/blitextDevelops and trains models for vision-language learning with decoupled language pre-training24
jshilong/gpt4roiTraining and deploying large language models on computer vision tasks using region-of-interest inputs517
deepseek-ai/deepseek-vlA multimodal AI model that enables real-world vision-language understanding applications2,145