CoBSAT

ML Model Benchmarker

Provides a benchmarking framework and dataset for evaluating the performance of large language models in text-to-image tasks

Implementation and dataset for paper "Can MLLMs Perform Text-to-Image In-Context Learning?"

GitHub

30 stars
1 watching
1 forks
Language: Jupyter Notebook
last commit: almost 2 years ago

Related projects:

RepositoryDescriptionStars
ys-zong/vl-iclA benchmarking suite for multimodal in-context learning models31
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
mshukor/evalign-iclEvaluating and improving large multimodal models through in-context learning21
junyangwang0410/amberAn LLM-free benchmark suite for evaluating MLLMs' hallucination capabilities in various tasks and dimensions98
freedomintelligence/mllm-benchEvaluates and compares the performance of multimodal large language models on various tasks56
ml-tooling/ml-workspaceAn all-in-one web-based IDE for machine learning and data science3,446
ailab-cvc/seed-benchA benchmark for evaluating large language models' ability to process multimodal input322
marklogic/ml-gradleAutomates tasks involving MarkLogic using Gradle72
trekhleb/machine-learning-experimentsAn interactive platform for exploring and comparing various machine learning algorithms and techniques using visualizations and example code.1,667
damo-nlp-sg/m3examA benchmark for evaluating large language models in multiple languages and formats93
mmaul/clmlA high-performance statistical machine learning library written in Common Lisp261
mmaul/clml.tutorialsTutorials and resources for learning Common Lisp Machine Learning with CLML.31
ardanlabs/training-aiProvides training materials and tools for building machine learning applications72
isekai-portal/link-context-learningAn implementation of a multimodal learning approach to improve language models' ability to recognize unseen images and understand novel concepts.91
kei500/liblinear-rubyProvides an interface to train and predict with machine learning models using LIBLINEAR83