GPT4RoI

Region-of-Interest Training

Training and deploying large language models on computer vision tasks using region-of-interest inputs

GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest

GitHub

517 stars
8 watching
25 forks
Language: Python
last commit: over 2 years ago
computer-visiongptllmmultimodalityroi

Related projects:

RepositoryDescriptionStars
pku-yuangroup/languagebindExtending pretraining models to handle multiple modalities by aligning language and video representations751
openai/finetune-transformer-lmThis project provides code and model for improving language understanding through generative pre-training using a transformer-based architecture.2,167
openai/lm-human-preferencesTraining methods and tools for fine-tuning language models using human preferences1,240
pku-yuangroup/moe-llavaA large vision-language model using a mixture-of-experts architecture to improve performance on multi-modal learning tasks2,023
megvii-research/tlcImproves image restoration performance by converting global operations to local ones during inference231
shawn-ieitsystems/yuan-1.0Large-scale language model with improved performance on NLP tasks through distributed training and efficient data processing591
yiren-jian/blitextDevelops and trains models for vision-language learning with decoupled language pre-training24
brightmart/xlnet_zhTrains a large Chinese language model on massive data and provides a pre-trained model for downstream tasks230
byungkwanlee/moaiImproves performance of vision language tasks by integrating computer vision capabilities into large language models314
ailab-cvc/gpt4toolsAn intelligent system that enables automatic control and utilization of visual foundation models to interact with images in conversational settings.762
kaiyangzhou/dassl.pytorchA PyTorch toolbox for supporting research and development of domain adaptation, generalization, and semi-supervised learning methods in computer vision.1,236
rdspring1/pytorch_gbw_lmTrains a large-scale PyTorch language model on the 1-Billion Word dataset123
microsoft/megatron-deepspeedResearch tool for training large transformer language models at scale1,926
pku-yuangroup/video-benchEvaluates and benchmarks large language models' video understanding capabilities121
ethanyanjiali/minchatgptThis project demonstrates the effectiveness of reinforcement learning from human feedback (RLHF) in improving small language models like GPT-2.214