ROLL-VideoQA
by noagarcia
PyTorch code for ROLL, a knowledge-based video story question answering model.
AI summary
VideoQA model
A PyTorch-based model for answering questions about videos based on unseen scenes and storylines
- stars
- 19
- forks
- 4
- watching
- 3
Similar projects
Found by comparing what the projects do, not just their names.
VQA model
A PyTorch implementation of visual question answering with multimodal representation learning
VQA system
PyTorch implementation of video question answering system based on TVQA dataset
Reading Comprehension Model
Implementing reading comprehension from Wikipedia questions to answer open-domain queries using PyTorch and SQuAD dataset
VQA Model Trainer
Implementations and tools for training and fine-tuning a visual question answering model based on the 2017 CVPR workshop winner's approach.
Video-language model
An efficient framework for end-to-end learning on image-text and video-text tasks
Movie QA system
This project explores question-answering in movies using various machine learning approaches.
VQA model framework
A software framework for training and deploying multimodal visual question answering models using compact bilinear pooling.
VQA system
An implementation of a VQA system using bottom-up attention, aiming to improve the efficiency and speed of visual question answering tasks.
VQA model
A Visual Question Answering model using a deeper LSTM and normalized CNN architecture.
VQA prompter
An implementation of a two-stage framework designed to prompt large language models with answer heuristics for knowledge-based visual question answering tasks.
Video conversational model
A video conversation model that generates meaningful conversations about videos using large vision and language models
Image QA model
This project provides code for training image question answering models using stacked attention networks and convolutional neural networks.
Medical image understanding toolkit
A medical visual question-answering dataset and toolkit for training models to understand medical images and instructions.
Video deblurring model
An open-source implementation of a deep learning model for video deblurring and motion estimation.
State vector simulator
A PyTorch-based simulator for quantum machine learning