MMVP

Visual model evaluation

An evaluation framework for multimodal language models' visual capabilities using image and question benchmarks.

GitHub

296 stars
10 watching
7 forks
Language: Python
last commit: over 2 years ago

Related projects:

RepositoryDescriptionStars
zhourax/vegaDevelops a multimodal task and dataset to assess vision-language models' ability to handle interleaved image-text inputs.33
modelscope/evalscopeA framework for efficiently evaluating and benchmarking large models308
yuweihao/mm-vetEvaluates the capabilities of large multimodal models using a set of diverse tasks and metrics274
chenllliang/mmevalproA benchmarking framework for evaluating Large Multimodal Models by providing rigorous metrics and an efficient evaluation pipeline.22
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
openbmb/viscpmA family of large multimodal models supporting multimodal conversational capabilities and text-to-image generation in multiple languages1,098
xverse-ai/xverse-v-13bA large multimodal model for visual question answering, trained on a dataset of 2.1B image-text pairs and 8.2M instruction sequences.78
huggingface/evaluateAn evaluation framework for machine learning models and datasets, providing standardized metrics and tools for comparing model performance.2,063
freedomintelligence/mllm-benchEvaluates and compares the performance of multimodal large language models on various tasks56
aifeg/benchlmmAn open-source benchmarking framework for evaluating cross-style visual capability of large multimodal models84
mlo-lab/muviA software framework for multi-view latent variable modeling with domain-informed structured sparsity27
yuliang-liu/monkeyAn end-to-end image captioning system that uses large multi-modal models and provides tools for training, inference, and demo usage.1,849
evolvinglmms-lab/lmms-evalTools and evaluation framework for accelerating the development of large multimodal models by providing an efficient way to assess their performance2,164
microsoft/mm-reactAn AI-powered system that leverages multimodal reasoning and action to analyze visual data and provide insights940
opengvlab/multi-modality-arenaAn evaluation platform for comparing multi-modality models on visual question-answering tasks478