AdaptiveAttention

Image Captioning Model

Adaptive attention mechanism for image captioning using visual sentinels

Implementation of "Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning"

GitHub

335 stars
13 watching
74 forks
Language: Jupyter Notebook
last commit: over 8 years ago
Linked from 2 awesome lists

attention-mechanismimage-captioningtorch

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
chapternewscu/image-captioning-with-semantic-attentionA deep learning model for generating image captions with semantic attention51
yunjey/show-attend-and-tellGenerates captions for images using an attention-based neural network907
contextualai/lensEnhances language models to generate text based on visual descriptions of images352
jazzsaxmafia/show_attend_and_tell.tensorflowA TensorFlow implementation of a neural caption generator using attention mechanisms.506
peteanderson80/bottom-up-attentionTrains a bottom-up attention model using Faster R-CNN and Visual Genome annotations for image captioning and VQA tasks1,438
lancopku/iaisThis project proposes a novel method for calibrating attention distributions in multimodal models to improve contextualized representations of image-text pairs.30
deeprnn/image_captioningThis implementation allows users to generate captions from images using a neural network model with visual attention.790
luoweizhou/vlpA project for pre-training models to support image captioning and question answering tasks.416
mikeswanson/jbwatchactivityindicatorGenerates Apple Watch activity indicator images for animation529
lukemelas/image-paragraph-captioningTrains image paragraph captioning models to generate diverse and accurate captions90
cshizhe/asg2capAn image caption generation model that uses abstract scene graphs to fine-grained control and generate captions200
fengyang0317/unsupervised_captioningAn unsupervised image captioning framework that allows generating captions from images without paired data.215
jiasenlu/hiecoattenvqaA framework for training Hierarchical Co-Attention models for Visual Question Answering using preprocessed data and a specific image model.349
xiadingz/video-caption.pytorchPyTorch implementation of video captioning, combining deep learning and computer vision techniques.402
yiwuzhong/sub-gcA PyTorch implementation of image captioning models via scene graph decomposition.96