GPT4RoI
Region-of-Interest Training
Training and deploying large language models on computer vision tasks using region-of-interest inputs
GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest
517 stars
8 watching
25 forks
Language: Python
last commit: over 2 years agocomputer-visiongptllmmultimodalityroi
Related projects:
| Repository | Description | Stars |
|---|---|---|
| Extending pretraining models to handle multiple modalities by aligning language and video representations | 751 | |
| This project provides code and model for improving language understanding through generative pre-training using a transformer-based architecture. | 2,167 | |
| Training methods and tools for fine-tuning language models using human preferences | 1,240 | |
| A large vision-language model using a mixture-of-experts architecture to improve performance on multi-modal learning tasks | 2,023 | |
| Improves image restoration performance by converting global operations to local ones during inference | 231 | |
| Large-scale language model with improved performance on NLP tasks through distributed training and efficient data processing | 591 | |
| Develops and trains models for vision-language learning with decoupled language pre-training | 24 | |
| Trains a large Chinese language model on massive data and provides a pre-trained model for downstream tasks | 230 | |
| Improves performance of vision language tasks by integrating computer vision capabilities into large language models | 314 | |
| An intelligent system that enables automatic control and utilization of visual foundation models to interact with images in conversational settings. | 762 | |
| A PyTorch toolbox for supporting research and development of domain adaptation, generalization, and semi-supervised learning methods in computer vision. | 1,236 | |
| Trains a large-scale PyTorch language model on the 1-Billion Word dataset | 123 | |
| Research tool for training large transformer language models at scale | 1,926 | |
| Evaluates and benchmarks large language models' video understanding capabilities | 121 | |
| This project demonstrates the effectiveness of reinforcement learning from human feedback (RLHF) in improving small language models like GPT-2. | 214 |