LLaVA-Plus-Codebase

Model trainer

A platform for training and deploying large language and vision models that can use tools to perform tasks

LLaVA-Plus: Large Language and Vision Assistants that Plug and Learn to Use Skills

GitHub

717 stars
12 watching
53 forks
Language: Python
last commit: over 2 years ago
agentlarge-language-modelslarge-multimodal-modelsmultimodal-large-language-modelstool-use

Related projects:

RepositoryDescriptionStars
wisconsinaivision/vip-llavaA system designed to enable large multimodal models to understand arbitrary visual prompts302
llava-vl/llava-interactive-demoAn all-in-one demo for interactive image processing and generation353
alibaba/conv-llavaThis project presents an optimization technique for large-scale image models to reduce computational requirements while maintaining performance.106
vpgtrans/vpgtransTransfers visual prompt generators across large language models to reduce training costs and enable customization of multimodal LLMs270
csuhan/onellmA framework for training and fine-tuning multimodal language models on various data types601
bobazooba/xllmA tool for training and fine-tuning large language models using advanced techniques387
openbmb/cpm-liveA live training platform for large-scale deep learning models, allowing community participation and collaboration in model development and deployment.511
flagai-open/aquila2Provides pre-trained language models and tools for fine-tuning and evaluation439
chendelong1999/polite-flamingoDevelops training methods to improve the politeness and natural flow of multi-modal Large Language Models63
yfzhang114/llava-alignDebiasing techniques to minimize hallucinations in large visual language models75
vhellendoorn/code-lmsA guide to using pre-trained large language models in source code analysis and generation1,789
mlpc-ucsd/blivaA multimodal LLM designed to handle text-rich visual questions270
vishaal27/sus-xThis is an open-source project that proposes a novel method to train large-scale vision-language models with minimal resources and no fine-tuning required.94
volcengine/vescaleA PyTorch-based framework for training large language models in parallel on multiple devices679
microsoft/llava-medA research project aimed at building large language and vision models for biomedical applications with capabilities comparable to GPT-4.1,622