densecap

Image describer

A deep learning framework for generating natural language descriptions of images by detecting objects and their attributes

Dense image captioning in Torch

GitHub

2k stars
68 watching
429 forks
Language: Jupyter Notebook
last commit: about 8 years ago
Linked from 2 awesome lists


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
jcjohnson/cnn-visThis project enables users to generate images using convolutional neural networks (CNNs) and visualize their activations.499
reinhardh/supplement_deep_decoderThis repository provides code for an image generating deep neural network designed to produce concise representations of images from few parameters.96
jaywongwang/densevideocaptioningAn implementation of a dense video captioning model with attention-based fusion and context gating149
vision-cair/chatcaptionerEnables automatic generation of descriptive text from images and videos based on user input.457
zhegan27/semantic_compositional_netsA deep learning framework providing a model architecture and training code for image captioning using semantic compositional networks70
ibm/max-image-caption-generatorAn image caption generation system utilizing machine learning models and deep neural networks.84
chapternewscu/image-captioning-with-semantic-attentionA deep learning model for generating image captions with semantic attention51
deeprnn/image_captioningThis implementation allows users to generate captions from images using a neural network model with visual attention.790
ucbdrive/dlaA software framework for deep learning-based image classification and segmentation tasks434
vision-cair/longvuAn artificial intelligence system designed to understand and describe long-form video content329
chxj1992/captcha_crackerAn image recognition system using a deep learning model to classify characters from verification codes189
jhcho99/coformerAn implementation of a deep learning model for grounding situation recognition in images45
zhujun98/semantic_segmentationImplementations of deep learning architectures for semantic segmentation of images in various datasets.6
codingjoe/django-picturesA Django package for responsive image handling using modern formats like AVIF and WebP251
contextualai/lensEnhances language models to generate text based on visual descriptions of images352