unsupervised_captioning

Image captioner

An unsupervised image captioning framework that allows generating captions from images without paired data.

Code for Unsupervised Image Captioning

GitHub

215 stars
7 watching
51 forks
Language: Python
last commit: over 3 years ago

Related projects:

RepositoryDescriptionStars
apple2373/chainer-captionAn image caption generation system using a neural network architecture with pre-trained models.64
deeprnn/image_captioningThis implementation allows users to generate captions from images using a neural network model with visual attention.790
chapternewscu/image-captioning-with-semantic-attentionA deep learning model for generating image captions with semantic attention51
luoweizhou/vlpA project for pre-training models to support image captioning and question answering tasks.416
cshizhe/asg2capAn image caption generation model that uses abstract scene graphs to fine-grained control and generate captions200
yiwuzhong/sub-gcA PyTorch implementation of image captioning models via scene graph decomposition.96
rmokady/clip_prefix_captionAn approach to image captioning that leverages the CLIP model and fine-tunes a language model without requiring additional supervision or object annotation.1,326
lukemelas/image-paragraph-captioningTrains image paragraph captioning models to generate diverse and accurate captions90
anonymousanoy/foheAutomates the process of generating multiple rewritten image captions by fine-tuning large vision-language models8
contextualai/lensEnhances language models to generate text based on visual descriptions of images352
ibm/max-image-caption-generatorAn image caption generation system utilizing machine learning models and deep neural networks.84
ttengwang/caption-anythingA tool generating descriptive captions from images with customizable controls and text styles.1,693
eladhoffer/captiongenA PyTorch-based tool for generating captions from images128
facebookresearch/cutlerAn unsupervised object detection and segmentation framework that can learn from image data alone, without requiring human annotations.954
jamespark3922/adv-infA method for generating and evaluating video captions using adversarial inference, trained on large datasets of text and multimedia features.34