MuVI

Multi-view modeler

A software framework for multi-view latent variable modeling with domain-informed structured sparsity

A multi-view latent variable model with domain-informed structured sparsity for integrating noisy feature sets.

GitHub

27 stars
5 watching
2 forks
Language: Python
last commit: over 1 year ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
tsb0601/mmvpAn evaluation framework for multimodal language models' visual capabilities using image and question benchmarks.296
yuliang-liu/monkeyAn end-to-end image captioning system that uses large multi-modal models and provides tools for training, inference, and demo usage.1,849
mcs07/molvsTool for validating and standardizing chemical structures to improve data quality and facilitate comparisons.163
vita-mllm/vitaA large multimodal language model designed to process and analyze video, image, text, and audio inputs in real-time.1,005
nvlabs/eagleDevelops high-resolution multimodal LLMs by combining vision encoders and various input resolutions549
yuweihao/mm-vetEvaluates the capabilities of large multimodal models using a set of diverse tasks and metrics274
opengvlab/multi-modality-arenaAn evaluation platform for comparing multi-modality models on visual question-answering tasks478
yfzhang114/slimeDevelops large multimodal models for high-resolution understanding and analysis of text, images, and other data types.143
xverse-ai/xverse-v-13bA large multimodal model for visual question answering, trained on a dataset of 2.1B image-text pairs and 8.2M instruction sequences.78
pku-yuangroup/moe-llavaA large vision-language model using a mixture-of-experts architecture to improve performance on multi-modal learning tasks2,023
zhourax/vegaDevelops a multimodal task and dataset to assess vision-language models' ability to handle interleaved image-text inputs.33
freedomintelligence/mllm-benchEvaluates and compares the performance of multimodal large language models on various tasks56
subho406/omninetAn implementation of a unified architecture for multi-modal multi-task learning using PyTorch.515
chenllliang/mmevalproA benchmarking framework for evaluating Large Multimodal Models by providing rigorous metrics and an efficient evaluation pipeline.22
mlpc-ucsd/blivaA multimodal LLM designed to handle text-rich visual questions270