SEED-Bench
by AILab-CVC
(CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions.
AI summary
Multimodal LLM test suite
A benchmark for evaluating large language models' ability to process multimodal input
- stars
- 322
- forks
- 13
- watching
- 4
Similar projects
Found by comparing what the projects do, not just their names.
Multimodal LLM
An implementation of a multimodal language model with capabilities for comprehension and generation
Multimodal benchmarking
Evaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.
Visual Model Benchmark
An open-source benchmarking framework for evaluating cross-style visual capability of large multimodal models
Multimodal model evaluator
Evaluating and improving large multimodal models through in-context learning
Model evaluator
Evaluates and compares the performance of multimodal large language models on various tasks
Multimodal LLM
An implementation of a multimodal language model using locality-enhanced projection techniques
OCR Benchmark
An evaluation benchmark for OCR capabilities in large multmodal models.
Cheminformatics toolkit
Automated machine learning protocols for cheminformatics using Python
LM Benchmark
A benchmark for evaluating large language models in multiple languages and formats
Learning benchmark
A benchmarking suite for multimodal in-context learning models
VQA model
A multimodal LLM designed to handle text-rich visual questions
Chart model trainer
Develops a large-scale dataset and benchmark for training multimodal chart understanding models using large language models.
LLM benchmarker
A benchmarking framework for large language models
cloud-cv/evalai1.8K
Benchmarking tool
A platform for comparing and evaluating AI and machine learning algorithms at scale
MLLM benchmark
An LLM-free benchmark suite for evaluating MLLMs' hallucination capabilities in various tasks and dimensions