Bongard-HOI

Visual Reasoning Benchmark

A benchmarking tool and software framework for evaluating few-shot visual reasoning capabilities in computer vision models.

[CVPR 2022 (oral)] Bongard-HOI for benchmarking few-shot visual reasoning

GitHub

64 stars
7 watching
7 forks
Language: Python
last commit: almost 4 years ago
cvpr2022few-shot-learningpytorchvisual-reasoning

Related projects:

RepositoryDescriptionStars
nvlabs/relvitA deep learning framework designed to improve visual reasoning capabilities by utilizing concepts and semantic relations.64
tianyi-lab/hallusionbenchAn image-context reasoning benchmark designed to challenge large vision-language models and help improve their accuracy259
nvlabs/bongard-logoGenerates synthetic Bongard problems with minimal human intervention.51
davidmascharka/tbd-netsAn open-source implementation of a deep learning model designed to improve the balance between performance and interpretability in visual reasoning tasks.348
ailab-cvc/seed-benchA benchmark for evaluating large language models' ability to process multimodal input322
rowanz/r2cAn open-source project providing PyTorch code and data for a deep learning model that enables visual commonsense reasoning.466
lxtgh/omg-segDevelops an end-to-end model for multiple visual perception and reasoning tasks using a single encoder, decoder, and large language model.1,336
yuqifan1117/hallucidoctorThis project provides tools and frameworks to mitigate hallucinatory toxicity in visual instruction data, allowing researchers to fine-tune MLLM models on specific datasets.41
nvlabs/prismerA deep learning framework for training multi-modal models with vision and language capabilities.1,299
lavi-lab/visual-tableA project that generates visual representations tailored for general visual reasoning, leveraging hierarchical scene descriptions and instance-level world knowledge.14
kunpengli1994/vsrnAn open-source PyTorch implementation of a visual semantic reasoning model for image-text matching294
nv-tlabs/stealDevelops a method to create high-quality training data from noisy labels in semantic segmentation tasks.478
bonlime/keras-deeplab-v3-plusAn implementation of Deeplabv3+ in Keras with pretrained weights and customization options for semantic image segmentation.1,360
hms-dbmi/vivA toolkit for interactive visualization of high-resolution bioimaging data.290
vcciv/blvdA large-scale 5D semantics benchmark for autonomous driving171