OmniBench
A project for tri-modal LLM benchmarking and instruction tuning.
AI summary
Multimodal benchmarking
Evaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.
- stars
- 15
- forks
- 2
- watching
- 0
Similar projects
Found by comparing what the projects do, not just their names.
Multimodal LLM test suite
A benchmark for evaluating large language models' ability to process multimodal input
OCR Benchmark
An evaluation benchmark for OCR capabilities in large multmodal models.
NLP model
A large language model designed for research and application in natural language processing tasks.
Multimodal LLM
A multi-modal large language model that integrates natural language and visual capabilities with fine-tuning for various tasks
Multimodal alignment model
Extending pretraining models to handle multiple modalities by aligning language and video representations
LLM benchmarker
A benchmarking framework for large language models
openbmb/viscpm1.1K
Multimodal Models
A family of large multimodal models supporting multimodal conversational capabilities and text-to-image generation in multiple languages
Multimodal LLM
A multi-modal language model that integrates image, video, audio, and text data to improve language understanding and generation
Multimodal LLM
An implementation of a multimodal language model with capabilities for comprehension and generation
Multi-modal ML framework
An implementation of a unified architecture for multi-modal multi-task learning using PyTorch.
Multimodal evaluation framework
Develops a multimodal task and dataset to assess vision-language models' ability to handle interleaved image-text inputs.
ML Model Benchmarker
Provides a benchmarking framework and dataset for evaluating the performance of large language models in text-to-image tasks
Image captioner
An end-to-end image captioning system that uses large multi-modal models and provides tools for training, inference, and demo usage.
LM Benchmark
A benchmark for evaluating large language models in multiple languages and formats
tsb0601/mmvp296
Visual model evaluation
An evaluation framework for multimodal language models' visual capabilities using image and question benchmarks.