Visual-Instruction-Tuning
by BAAI-DCAI
SVIT: Scaling up Visual Instruction Tuning
AI summary
Visual Instruction Tuning
A dataset and model designed to scale visual instruction tuning using language-only GPT-4 models.
- stars
- 164
- forks
- 4
- watching
- 5
Similar projects
Found by comparing what the projects do, not just their names.
Visual Instruction Model
A project that extends large language models by integrating an additional region-level vision encoder to improve visual instruction tuning.
Visual Instruction Tuning Tool
A tool for generating and evaluating multimodal Large Language Models with visual instruction tuning capabilities
Visual Instruction Tuning
An open-source project that enhances visual instruction tuning for text-rich image understanding by integrating GPT-4 models with multimodal datasets.
Instruction generator
Creating synthetic visual reasoning instructions to improve the performance of large language models on image-related tasks
language model tuning
Training methods and tools for fine-tuning language models using human preferences
Visual Instruction Toolkit
A method and toolkit for fine-tuning large language models to perform visual instruction tasks in multiple languages.
Instruction dataset generator
Autonomously generates high-quality image-text instruction fine-tuning datasets
Region-of-Interest Training
Training and deploying large language models on computer vision tasks using region-of-interest inputs
Instruction dataset
A multimodal benchmark dataset designed to evaluate the performance of vision-language foundation models through instruction tuning.
Safety fine-tuner
Improves safety and helpfulness of large language models by fine-tuning them using safety-critical tasks
Performance tuning guide
A tuning guide for optimizing the performance of a network intrusion prevention system
pevma/septun204
Performance Tuning Guide
A guide to tuning Suricata for maximum performance in network intrusion detection systems
Visual guidance
This project presents a new approach to fine-grained visual understanding using pixel-wise mask regions in language instructions
Network tuner
Automates the search for optimal neural network configurations in deep learning applications
VLIT model toolkit
A comprehensive dataset and evaluation framework for Vision-Language Instruction Tuning models