Bongard-HOI
by NVlabs
[CVPR 2022 (oral)] Bongard-HOI for benchmarking few-shot visual reasoning
AI summary
Visual Reasoning Benchmark
A benchmarking tool and software framework for evaluating few-shot visual reasoning capabilities in computer vision models.
- stars
- 64
- forks
- 7
- watching
- 7
Similar projects
Found by comparing what the projects do, not just their names.
Visual reasoning tool
A deep learning framework designed to improve visual reasoning capabilities by utilizing concepts and semantic relations.
Benchmark
An image-context reasoning benchmark designed to challenge large vision-language models and help improve their accuracy
Problem generator
Generates synthetic Bongard problems with minimal human intervention.
Visual Reasoning Model
An open-source implementation of a deep learning model designed to improve the balance between performance and interpretability in visual reasoning tasks.
Multimodal LLM test suite
A benchmark for evaluating large language models' ability to process multimodal input
rowanz/r2c466
Visual Reasoning Model
An open-source project providing PyTorch code and data for a deep learning model that enables visual commonsense reasoning.
lxtgh/omg-seg1.3K
Visual Model
Develops an end-to-end model for multiple visual perception and reasoning tasks using a single encoder, decoder, and large language model.
Data processing framework
This project provides tools and frameworks to mitigate hallucinatory toxicity in visual instruction data, allowing researchers to fine-tune MLLM models on specific datasets.
nvlabs/prismer1.3K
Vision-Language Model
A deep learning framework for training multi-modal models with vision and language capabilities.
Visual representation generator
A project that generates visual representations tailored for general visual reasoning, leveraging hierarchical scene descriptions and instance-level world knowledge.
Visual matcher
An open-source PyTorch implementation of a visual semantic reasoning model for image-text matching
Label correction tool
Develops a method to create high-quality training data from noisy labels in semantic segmentation tasks.
Semantic image segementation model
An implementation of Deeplabv3+ in Keras with pretrained weights and customization options for semantic image segmentation.
hms-dbmi/viv290
Visualization toolkit
A toolkit for interactive visualization of high-resolution bioimaging data.
vcciv/blvd171
Autonomous driving dataset
A large-scale 5D semantics benchmark for autonomous driving