zsgnet-pytorch
by TheShadow29
Official implementation of ICCV19 oral paper Zero-Shot grounding of Objects from Natural Language Queries (https://arxiv.org/abs/1908.07129)
AI summary
Object groundings model
An implementation of a computer vision model that grounds objects in images using natural language queries.
- stars
- 69
- forks
- 12
- watching
- 4
Similar projects
Found by comparing what the projects do, not just their names.
Image grounding model
An implementation of a deep learning model for grounding situation recognition in images
Computer Vision Toolkit
A PyTorch toolbox for supporting research and development of domain adaptation, generalization, and semi-supervised learning methods in computer vision.
microsoft/som1.2K
Image marking tool
Enables visual grounding in large language models by overlaying spatial and speakable marks on images
Deep learning toolkit
A Python framework for building deep learning models with optimized encoding layers and batch normalization.
Multimodal conversational model
An end-to-end trained model capable of generating natural language responses integrated with object segmentation masks for interactive visual conversations
Semantic segmenation model
Re-implementation of a deep learning model for semantic segmentation using PyTorch.
Image Description Model
Trains multilingual image description models using neural sequence models and extracts hidden features from trained models.
Semantic Segmentation Model
An implementation of a deep learning model using PyTorch for semantic segmentation tasks.
RL Agent Trainer
Trains an RL agent to execute natural language instructions in a 3D environment using a combination of A3C and gated attention mechanisms.
Language model
An implementation of DeepMind's Relational Recurrent Neural Networks (Santoro et al. 2018) in PyTorch for word language modeling
hszhao/pspnet1.6K
Segmentation Model
A PyTorch implementation of a deep learning model for semantic image segmentation
Vision-Language Model Framework
Implementing a unified modal learning framework for generative vision-language models
Deep Learning Model
A deep learning model implementation of the DeepLab ResNet architecture for image segmentation tasks.
Point cloud modeler
A PyTorch framework for building and training deep learning models on point clouds.
Point Cloud Model
This is an implementation of the PointNet algorithm in PyTorch for 3D point cloud classification and segmentation tasks.