TVQA
by jayleicn
[EMNLP 2018] PyTorch code for TVQA: Localized, Compositional Video Question Answering
AI summary
VQA system
PyTorch implementation of video question answering system based on TVQA dataset
- stars
- 172
- forks
- 32
- watching
- 9
Similar projects
Found by comparing what the projects do, not just their names.
VQA model
A PyTorch implementation of visual question answering with multimodal representation learning
Video-language model
An efficient framework for end-to-end learning on image-text and video-text tasks
VQA Model Trainer
Implementations and tools for training and fine-tuning a visual question answering model based on the 2017 CVPR workshop winner's approach.
VideoQA model
A PyTorch-based model for answering questions about videos based on unseen scenes and storylines
QNA architecture
A PyTorch implementation of an improved question answering architecture with dynamic memory networks and attention mechanisms
Reading Comprehension Model
Implementing reading comprehension from Wikipedia questions to answer open-domain queries using PyTorch and SQuAD dataset
VQA trainer
Tools and scripts for training and evaluating a visual question answering model using transfer learning from an external data source.
State vector simulator
A PyTorch-based simulator for quantum machine learning
Video captioner
PyTorch implementation of video captioning, combining deep learning and computer vision techniques.
VQA system
An implementation of a VQA system using bottom-up attention, aiming to improve the efficiency and speed of visual question answering tasks.
Semantic Segmentation Model
A PyTorch implementation of a real-time semantic segmentation model using ENet architecture
VQA model
A Visual Question Answering model using a deeper LSTM and normalized CNN architecture.
NMN
A PyTorch implementation of Neural Module Networks for Visual Question Answering
PNASNet model
PyTorch implementation of PNASNet-5 architecture
GA trainer
Trains PyTorch models using a genetic algorithm