SEED-Bench

Multimodal LLM test suite

A benchmark for evaluating large language models' ability to process multimodal input

(CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions.

GitHub

322 stars
4 watching
13 forks
Language: Python
last commit: about 2 years ago

Related projects:

RepositoryDescriptionStars
ailab-cvc/seedAn implementation of a multimodal language model with capabilities for comprehension and generation585
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
aifeg/benchlmmAn open-source benchmarking framework for evaluating cross-style visual capability of large multimodal models84
mshukor/evalign-iclEvaluating and improving large multimodal models through in-context learning21
freedomintelligence/mllm-benchEvaluates and compares the performance of multimodal large language models on various tasks56
khanrc/honeybeeAn implementation of a multimodal language model using locality-enhanced projection techniques435
yuliang-liu/multimodalocrAn evaluation benchmark for OCR capabilities in large multmodal models.484
jvalegre/robertAutomated machine learning protocols for cheminformatics using Python39
damo-nlp-sg/m3examA benchmark for evaluating large language models in multiple languages and formats93
ys-zong/vl-iclA benchmarking suite for multimodal in-context learning models31
mlpc-ucsd/blivaA multimodal LLM designed to handle text-rich visual questions270
fuxiaoliu/mmcDevelops a large-scale dataset and benchmark for training multimodal chart understanding models using large language models.87
qcri/llmebenchA benchmarking framework for large language models81
cloud-cv/evalaiA platform for comparing and evaluating AI and machine learning algorithms at scale1,779
junyangwang0410/amberAn LLM-free benchmark suite for evaluating MLLMs' hallucination capabilities in various tasks and dimensions98