MMVP
by tsb0601
AI summary
Visual model evaluation
An evaluation framework for multimodal language models' visual capabilities using image and question benchmarks.
- stars
- 296
- forks
- 7
- watching
- 10
Similar projects
Found by comparing what the projects do, not just their names.
Multimodal evaluation framework
Develops a multimodal task and dataset to assess vision-language models' ability to handle interleaved image-text inputs.
Model evaluator
A framework for efficiently evaluating and benchmarking large models
Model evaluator
Evaluates the capabilities of large multimodal models using a set of diverse tasks and metrics
Model Evaluator
A benchmarking framework for evaluating Large Multimodal Models by providing rigorous metrics and an efficient evaluation pipeline.
Multimodal benchmarking
Evaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.
openbmb/viscpm1.1K
Multimodal Models
A family of large multimodal models supporting multimodal conversational capabilities and text-to-image generation in multiple languages
Multimodal model
A large multimodal model for visual question answering, trained on a dataset of 2.1B image-text pairs and 8.2M instruction sequences.
Model Evaluator
An evaluation framework for machine learning models and datasets, providing standardized metrics and tools for comparing model performance.
Model evaluator
Evaluates and compares the performance of multimodal large language models on various tasks
Visual Model Benchmark
An open-source benchmarking framework for evaluating cross-style visual capability of large multimodal models
Multi-view modeler
A software framework for multi-view latent variable modeling with domain-informed structured sparsity
Image captioner
An end-to-end image captioning system that uses large multi-modal models and provides tools for training, inference, and demo usage.
Model evaluation toolkit
Tools and evaluation framework for accelerating the development of large multimodal models by providing an efficient way to assess their performance
Visual analyzer
An AI-powered system that leverages multimodal reasoning and action to analyze visual data and provide insights
Model arena
An evaluation platform for comparing multi-modality models on visual question-answering tasks