GPT4RoI
by jshilong
GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest
AI summary
Region-of-Interest Training
Training and deploying large language models on computer vision tasks using region-of-interest inputs
- stars
- 517
- forks
- 25
- watching
- 8
Similar projects
Found by comparing what the projects do, not just their names.
Multimodal alignment model
Extending pretraining models to handle multiple modalities by aligning language and video representations
Language model trainer
This project provides code and model for improving language understanding through generative pre-training using a transformer-based architecture.
language model tuning
Training methods and tools for fine-tuning language models using human preferences
Mixture of Experts Model
A large vision-language model using a mixture-of-experts architecture to improve performance on multi-modal learning tasks
Localizer
Improves image restoration performance by converting global operations to local ones during inference
Language Model
Large-scale language model with improved performance on NLP tasks through distributed training and efficient data processing
Vision-Language Learning Model
Develops and trains models for vision-language learning with decoupled language pre-training
Chinese language model
Trains a large Chinese language model on massive data and provides a pre-trained model for downstream tasks
Vision Language Integrator
Improves performance of vision language tasks by integrating computer vision capabilities into large language models
Conversational image interface
An intelligent system that enables automatic control and utilization of visual foundation models to interact with images in conversational settings.
Computer Vision Toolkit
A PyTorch toolbox for supporting research and development of domain adaptation, generalization, and semi-supervised learning methods in computer vision.
Language Model
Trains a large-scale PyTorch language model on the 1-Billion Word dataset
Transformer trainer
Research tool for training large transformer language models at scale
Video benchmarking toolkit
Evaluates and benchmarks large language models' video understanding capabilities
Model alignment
This project demonstrates the effectiveness of reinforcement learning from human feedback (RLHF) in improving small language models like GPT-2.