MultimodalOCR

OCR Benchmark

An evaluation benchmark for OCR capabilities in large multmodal models.

On the Hidden Mystery of OCR in Large Multimodal Models (OCRBench)

GitHub

484 stars
15 watching
32 forks
Language: Python
last commit: almost 2 years ago

Related projects:

RepositoryDescriptionStars
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
yuliang-liu/monkeyAn end-to-end image captioning system that uses large multi-modal models and provides tools for training, inference, and demo usage.1,849
ailab-cvc/seed-benchA benchmark for evaluating large language models' ability to process multimodal input322
mshukor/evalign-iclEvaluating and improving large multimodal models through in-context learning21
ys-zong/vl-iclA benchmarking suite for multimodal in-context learning models31
felixgithub2017/mmcuMeasures the understanding of massive multitask Chinese datasets using large language models87
oeg-upm/lubm4obdaEvaluates Ontology-Based Data Access systems with inference and meta knowledge benchmarking4
openml/automlbenchmarkA framework for evaluating and comparing machine learning pipelines and neural architectures.413
qcri/llmebenchA benchmarking framework for large language models81
aifeg/benchlmmAn open-source benchmarking framework for evaluating cross-style visual capability of large multimodal models84
yuweihao/mm-vetEvaluates the capabilities of large multimodal models using a set of diverse tasks and metrics274
uw-madison-lee-lab/cobsatProvides a benchmarking framework and dataset for evaluating the performance of large language models in text-to-image tasks30
pkunlp-icler/pca-evalAn open-source benchmark and evaluation tool for assessing multimodal large language models' performance in embodied decision-making tasks99
chenllliang/mmevalproA benchmarking framework for evaluating Large Multimodal Models by providing rigorous metrics and an efficient evaluation pipeline.22
fuxiaoliu/mmcDevelops a large-scale dataset and benchmark for training multimodal chart understanding models using large language models.87