VL-ICL

Learning benchmark

A benchmarking suite for multimodal in-context learning models

Code for paper: VL-ICL Bench: The Devil in the Details of Benchmarking Multimodal In-Context Learning

GitHub

31 stars
1 watching
2 forks
Language: Python
last commit: over 2 years ago

Related projects:

RepositoryDescriptionStars
mshukor/evalign-iclEvaluating and improving large multimodal models through in-context learning21
uw-madison-lee-lab/cobsatProvides a benchmarking framework and dataset for evaluating the performance of large language models in text-to-image tasks30
ys-zong/vlguardImproves safety and helpfulness of large language models by fine-tuning them using safety-critical tasks47
ailab-cvc/seed-benchA benchmark for evaluating large language models' ability to process multimodal input322
yg-smile/rl_vvc_datasetA collection of benchmarks and implementations for testing reinforcement learning-based Volt-VAR control algorithms20
haozhezhao/micDevelops a multimodal vision-language model to enable machines to understand complex relationships between instructions and images in various tasks.337
yuliang-liu/multimodalocrAn evaluation benchmark for OCR capabilities in large multmodal models.484
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
ydli-ai/cslA large-scale dataset for natural language processing tasks focused on Chinese scientific literature, providing tools and benchmarks for NLP research.582
cloud-cv/evalaiA platform for comparing and evaluating AI and machine learning algorithms at scale1,779
rll-research/url_benchmarkA benchmark suite for unsupervised reinforcement learning agents, providing pre-trained models and scripts for testing and fine-tuning agent performance.335
scicloj/scicloj.ml.clj-djlProvides pre-trained machine learning models for natural language processing tasks using Clojure and the clj-djl framework.0
yiren-jian/blitextDevelops and trains models for vision-language learning with decoupled language pre-training24
jiutian-vl/jiutian-lionThis project integrates visual knowledge into large language models to improve their capabilities and reduce hallucinations.124
pku-yuangroup/languagebindExtending pretraining models to handle multiple modalities by aligning language and video representations751