image-captioning-with-semantic-attention

Image captioning model

A deep learning model for generating image captions with semantic attention

GitHub

51 stars
6 watching
17 forks
Language: Jupyter Notebook
last commit: almost 10 years ago

Related projects:

RepositoryDescriptionStars
zhegan27/semantic_compositional_netsA deep learning framework providing a model architecture and training code for image captioning using semantic compositional networks70
deeprnn/image_captioningThis implementation allows users to generate captions from images using a neural network model with visual attention.790
jiasenlu/adaptiveattentionAdaptive attention mechanism for image captioning using visual sentinels335
cshizhe/asg2capAn image caption generation model that uses abstract scene graphs to fine-grained control and generate captions200
fengyang0317/unsupervised_captioningAn unsupervised image captioning framework that allows generating captions from images without paired data.215
contextualai/lensEnhances language models to generate text based on visual descriptions of images352
ibm/max-image-caption-generatorAn image caption generation system utilizing machine learning models and deep neural networks.84
zhujun98/semantic_segmentationImplementations of deep learning architectures for semantic segmentation of images in various datasets.6
rmokady/clip_prefix_captionAn approach to image captioning that leverages the CLIP model and fine-tunes a language model without requiring additional supervision or object annotation.1,326
yunjey/show-attend-and-tellGenerates captions for images using an attention-based neural network907
apple2373/chainer-captionAn image caption generation system using a neural network architecture with pre-trained models.64
jaywongwang/densevideocaptioningAn implementation of a dense video captioning model with attention-based fusion and context gating149
speedinghzl/ccnetAn implementation of a deep learning model for semantic segmentation using a novel attention mechanism to capture long-range dependencies in images.1,432
zhengpeng7/birefnetAn open-source implementation of an image segmentation model that combines background removal and object detection capabilities.1,484
lukemelas/image-paragraph-captioningTrains image paragraph captioning models to generate diverse and accurate captions90