MoE-LLaVA
Mixture-of-Experts for Large Vision-Language Models
AI summary
Mixture of Experts Model
A large vision-language model using a mixture-of-experts architecture to improve performance on multi-modal learning tasks
- stars
- 2K
- forks
- 126
- watching
- 24
Similar projects
Found by comparing what the projects do, not just their names.
Mixture-of-Experts Model
Developed by XVERSE Technology Inc. as a multilingual large language model with a unique mixture-of-experts architecture and fine-tuned for various tasks such as conversation, question answering, and natural language understanding.
Mixture-of-experts model
A method for scaling multimodal large language models by combining multiple experts and fine-tuning them together
Vision Language Integrator
Improves performance of vision language tasks by integrating computer vision capabilities into large language models
MoE model
A high-performance mixture-of-experts model with innovative training techniques for language processing tasks
Model debiasing
Debiasing techniques to minimize hallucinations in large visual language models
Multilingual Model
Develops and publishes large multilingual language models with advanced mixing-of-experts architecture.
Multimodal alignment model
Extending pretraining models to handle multiple modalities by aligning language and video representations
Region-of-Interest Training
Training and deploying large language models on computer vision tasks using region-of-interest inputs
Efficient LLM
A large language model with improved efficiency and performance compared to similar models
Multimodal LLM
A multi-modal large language model that integrates natural language and visual capabilities with fine-tuning for various tasks
Mixtral model developer
Develops and releases Mixtral-based models for natural language processing tasks with a focus on Chinese text generation and understanding
Model optimizer
This project presents an optimization technique for large-scale image models to reduce computational requirements while maintaining performance.
Visual encoder
A vision-language model that uses a query transformer to encode images as visual tokens and allows flexible choice of the number of visual tokens.
Language Model
A high-performance language model designed to excel in tasks like natural language understanding, mathematical computation, and code generation
Model trainer
A platform for training and deploying large language and vision models that can use tools to perform tasks