MMBench
by open-compass
Official Repo of "MMBench: Is Your Multi-modal Model an All-around Player?"
AI summary
Multi-modal model evaluation suite
A collection of benchmarks to evaluate the multi-modal understanding capability of large vision language models.
- stars
- 168
- forks
- 10
- watching
- 3
Similar projects
Found by comparing what the projects do, not just their names.
Evaluation framework
An evaluation toolkit for large vision-language models
Multimodal benchmarking
Evaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.
openbmb/viscpm1.1K
Multimodal Models
A family of large multimodal models supporting multimodal conversational capabilities and text-to-image generation in multiple languages
Legal model evaluator
Evaluates the legal knowledge of large language models using a custom benchmarking framework.
Model arena
An evaluation platform for comparing multi-modality models on visual question-answering tasks
Multi-modal language model
An empirical study aiming to develop a large language model capable of effectively integrating multiple input modalities
3D dataset
An open-source software project providing a comprehensive 3D instruction-following dataset with multi-modal prompts for training large language models.
Chart model trainer
Develops a large-scale dataset and benchmark for training multimodal chart understanding models using large language models.
tsb0601/mmvp296
Visual model evaluation
An evaluation framework for multimodal language models' visual capabilities using image and question benchmarks.
Model robustness tester
A benchmarking framework designed to evaluate the robustness of large multimodal models against common corruption scenarios
Neural analysis tool
A MATLAB package for modeling and analyzing multivariate neural responses to dynamic stimuli.
Human model toolkit
Provides a modular framework and tools for working with 3D human parametric models in computer vision and graphics
Action understanding tool
An open-source toolbox for action understanding from video data using PyTorch.
Multimodal LLM test suite
A benchmark for evaluating large language models' ability to process multimodal input
Multimodal LLM
A multi-modal large language model that integrates natural language and visual capabilities with fine-tuning for various tasks