OmniBench

Multimodal benchmarking

Evaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.

A project for tri-modal LLM benchmarking and instruction tuning.

GitHub

15 stars
0 watching
2 forks
Language: Python
last commit: almost 2 years ago

Related projects:

RepositoryDescriptionStars
ailab-cvc/seed-benchA benchmark for evaluating large language models' ability to process multimodal input322
yuliang-liu/multimodalocrAn evaluation benchmark for OCR capabilities in large multmodal models.484
multimodal-art-projection/map-neoA large language model designed for research and application in natural language processing tasks.887
pleisto/yuren-baichuan-7bA multi-modal large language model that integrates natural language and visual capabilities with fine-tuning for various tasks73
pku-yuangroup/languagebindExtending pretraining models to handle multiple modalities by aligning language and video representations751
qcri/llmebenchA benchmarking framework for large language models81
openbmb/viscpmA family of large multimodal models supporting multimodal conversational capabilities and text-to-image generation in multiple languages1,098
lyuchenyang/macaw-llmA multi-modal language model that integrates image, video, audio, and text data to improve language understanding and generation1,568
ailab-cvc/seedAn implementation of a multimodal language model with capabilities for comprehension and generation585
subho406/omninetAn implementation of a unified architecture for multi-modal multi-task learning using PyTorch.515
zhourax/vegaDevelops a multimodal task and dataset to assess vision-language models' ability to handle interleaved image-text inputs.33
uw-madison-lee-lab/cobsatProvides a benchmarking framework and dataset for evaluating the performance of large language models in text-to-image tasks30
yuliang-liu/monkeyAn end-to-end image captioning system that uses large multi-modal models and provides tools for training, inference, and demo usage.1,849
damo-nlp-sg/m3examA benchmark for evaluating large language models in multiple languages and formats93
tsb0601/mmvpAn evaluation framework for multimodal language models' visual capabilities using image and question benchmarks.296