Semantic_Compositional_Nets

Image Captioning Model

A deep learning framework providing a model architecture and training code for image captioning using semantic compositional networks

The Theano code for the CVPR 2017 paper "Semantic Compositional Networks for Visual Captioning"

GitHub

70 stars
6 watching
24 forks
Language: Python
last commit: over 8 years ago

Related projects:

RepositoryDescriptionStars
chapternewscu/image-captioning-with-semantic-attentionA deep learning model for generating image captions with semantic attention51
deeprnn/image_captioningThis implementation allows users to generate captions from images using a neural network model with visual attention.790
speedinghzl/ccnetAn implementation of a deep learning model for semantic segmentation using a novel attention mechanism to capture long-range dependencies in images.1,432
zhujun98/semantic_segmentationImplementations of deep learning architectures for semantic segmentation of images in various datasets.6
preritj/segmentationDeep learning models for semantic segmentation of images101
nv-tlabs/gscnnThis code implements a neural network architecture designed to perform semantic segmentation in computer vision tasks.920
hszhao/pspnetA PyTorch implementation of a deep learning model for semantic image segmentation1,598
zhengpeng7/birefnetAn open-source implementation of an image segmentation model that combines background removal and object detection capabilities.1,484
jaywongwang/densevideocaptioningAn implementation of a dense video captioning model with attention-based fusion and context gating149
deepcs233/visual-cotA framework for training multi-modal language models with a focus on visual inputs and providing interpretable thoughts.162
pathak22/context-encoderUnsupervised feature learning by image inpainting using Generative Adversarial Networks (GANs)887
k3nt0w/fcn_via_kerasA Python implementation of a deep neural network architecture for semantic image segmentation48
zhegan27/convsentTrains an autoencoder to learn generic sentence representations using convolutional neural networks34
homles11/igcv3An implementation of an efficient deep neural network architecture189
cshizhe/asg2capAn image caption generation model that uses abstract scene graphs to fine-grained control and generate captions200