LLaVA-Interactive-Demo
by LLaVA-VL
LLaVA-Interactive-Demo
AI summary
Image processor
An all-in-one demo for interactive image processing and generation
- stars
- 353
- forks
- 27
- watching
- 16
Similar projects
Found by comparing what the projects do, not just their names.
Video image processor
An image-based language model that uses large language models to generate visual and text features from videos
Visual Prompt Model
A system designed to enable large multimodal models to understand arbitrary visual prompts
Model trainer
A platform for training and deploying large language and vision models that can use tools to perform tasks
Multimodal LLM
An implementation of a multimodal language model with capabilities for comprehension and generation
VQA model
A multimodal LLM designed to handle text-rich visual questions
Model optimizer
This project presents an optimization technique for large-scale image models to reduce computational requirements while maintaining performance.
Multimodal processor
A large multimodal language model designed to process and analyze video, image, text, and audio inputs in real-time.
nvlabs/eagle549
Multimodal model builder
Develops high-resolution multimodal LLMs by combining vision encoders and various input resolutions
Image processor
A system for scaling large language models to process and understand visual information from multiple images efficiently.
Image editing assistant
An implementation of a multimodal generation assistant using large language models and various image editing techniques.
LLM library
A Common Lisp port of a Large Language Model (LLM) implementation
Image processing library
A library of fast computer vision algorithms implemented in C++ for speed, operating over numpy arrays.
Visual Model
Develops a multimodal Chinese language model with visual capabilities
Image processor
A Lua binding for a fast image processing library with low memory needs.
libav/libav1.1K
Multimedia processor
A collection of libraries and tools for processing multimedia content