MultiInstruct

Instruction dataset

A multimodal benchmark dataset designed to evaluate the performance of vision-language foundation models through instruction tuning.

MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

GitHub

134 stars
7 watching
5 forks
Language: Python
last commit: over 3 years ago

Related projects:

RepositoryDescriptionStars
flagopen/flaginstructA collection of diverse instruction corpora for improving the development and tuning of Chinese Language Models173
x2fd/lvis-instruct4vA dataset of fine-grained visual instructions generated by prompting a large language model with images from another dataset131
pvit-official/pvitA project that extends large language models by integrating an additional region-level vision encoder to improve visual instruction tuning.37
salt-nlp/llavarAn open-source project that enhances visual instruction tuning for text-rich image understanding by integrating GPT-4 models with multimodal datasets.259
mbzuai-nlp/bactrian-xA collection of multilingual language models trained on a dataset of instructions and responses in various languages.94
opendatalab/vigcAutonomously generates high-quality image-text instruction fine-tuning datasets91
orhonovich/unnatural-instructionsA collection of automatically generated instructions for training language models.176
xuefuzhao/instructionwildCreating a large-scale user-based instruction dataset for natural language processing research and development455
baai-dcai/visual-instruction-tuningA dataset and model designed to scale visual instruction tuning using language-only GPT-4 models.164
zjunlp/mol-instructionsA dataset and tools package designed to support the training and evaluation of large language models for molecular biology tasks255
philipperemy/timitA collection of acoustic and phonetic speech data designed for training and evaluating automatic speech recognition systems297
michael-wzhu/promptcblueA large-scale instruction-tuning dataset for multi-task and few-shot learning in the medical domain328
dcdmllm/cheetahA large language model designed to understand and generate instructions with accompanying visual content360
rucaibox/comvintCreating synthetic visual reasoning instructions to improve the performance of large language models on image-related tasks18
aidc-ai/parrotA method and toolkit for fine-tuning large language models to perform visual instruction tasks in multiple languages.34