unsupervised_captioning
by fengyang0317
Code for Unsupervised Image Captioning
AI summary
Image captioner
An unsupervised image captioning framework that allows generating captions from images without paired data.
- stars
- 215
- forks
- 51
- watching
- 7
Similar projects
Found by comparing what the projects do, not just their names.
Image Captioner
An image caption generation system using a neural network architecture with pre-trained models.
Image captioning system
This implementation allows users to generate captions from images using a neural network model with visual attention.
Image captioning model
A deep learning model for generating image captions with semantic attention
Image Captioner
A project for pre-training models to support image captioning and question answering tasks.
Image captioning model
An image caption generation model that uses abstract scene graphs to fine-grained control and generate captions
Image Captioner
A PyTorch implementation of image captioning models via scene graph decomposition.
Image captioning method
An approach to image captioning that leverages the CLIP model and fine-tunes a language model without requiring additional supervision or object annotation.
Caption generator
Trains image paragraph captioning models to generate diverse and accurate captions
Caption rewriting
Automates the process of generating multiple rewritten image captions by fine-tuning large vision-language models
Image captioner
Enhances language models to generate text based on visual descriptions of images
Image generator
An image caption generation system utilizing machine learning models and deep neural networks.
Captioner
A tool generating descriptive captions from images with customizable controls and text styles.
Image caption generator
A PyTorch-based tool for generating captions from images
Image segmentation framework
An unsupervised object detection and segmentation framework that can learn from image data alone, without requiring human annotations.
Caption generation
A method for generating and evaluating video captions using adversarial inference, trained on large datasets of text and multimedia features.