LVIS-INSTRUCT4V

Visual Instructions

A dataset of fine-grained visual instructions generated by prompting a large language model with images from another dataset

GitHub

131 stars
3 watching
0 forks
last commit: over 2 years ago

Related projects:

RepositoryDescriptionStars
xuefuzhao/instructionwildCreating a large-scale user-based instruction dataset for natural language processing research and development455
flagopen/flaginstructA collection of diverse instruction corpora for improving the development and tuning of Chinese Language Models173
vt-nlp/multiinstructA multimodal benchmark dataset designed to evaluate the performance of vision-language foundation models through instruction tuning.134
orhonovich/unnatural-instructionsA collection of automatically generated instructions for training language models.176
jy0205/lavitA unified framework for training large language models to understand and generate visual content544
rucaibox/comvintCreating synthetic visual reasoning instructions to improve the performance of large language models on image-related tasks18
salt-nlp/llavarAn open-source project that enhances visual instruction tuning for text-rich image understanding by integrating GPT-4 models with multimodal datasets.259
ffxsam/vue-typescript-cookbookA cookbook and resource guide for developers learning Vue.js with TypeScript273
pvit-official/pvitA project that extends large language models by integrating an additional region-level vision encoder to improve visual instruction tuning.37
ncsoft/cap2qaA dataset and implementation of a method to generate instructions based on visual data5
aidc-ai/parrotA method and toolkit for fine-tuning large language models to perform visual instruction tasks in multiple languages.34
vefstathiou/so_word2vecThis is a word embedding model trained on Stack Overflow posts for use in natural language processing tasks.40
alexcode/vue2visProvides Vue.js 2 adapters for Vis.js libraries217
freedomintelligence/allavaA collection of datasets and models designed to support the training of lite vision-language models.249
vlf-silkie/vlfeedbackAn annotated preference dataset and training framework for improving large vision language models.88