TDD

Video descriptor extractor

A tool for extracting features from videos using deep convolutional descriptors

Trajectory-pooled Deep-Convolutional Descriptors

GitHub

104 stars
15 watching
75 forks
Language: Matlab
last commit: about 9 years ago
Linked from 1 awesome list

action-recognitioncaffe

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
wanglimin/untrimmednetA system for recognizing and detecting actions in untrimmed videos using a weakly supervised learning approach.162
damo-nlp-sg/vcdAn approach to reduce object hallucinations in large vision-language models by contrasting output distributions derived from original and distorted visual inputs222
csdms-contrib/dreich_algorithmA C++ algorithm for extracting channel networks from high resolution topographic data1
reedscot/cvpr2016A system for learning deep representations of fine-grained visual descriptions from images336
neil-wu/swiftdumpA tool that extracts information about Swift objects from Mach-O files.400
xuchaoxi/video-cnn-featExtracts CNN features from video frames using pre-trained MXNet models31
wangboml/bp_features_extractionA Matlab program for extracting features from three physiological signals (PPG, ECG, and BP) collected in synchronization.44
huaizhengzhang/awsome-deep-learning-for-video-analysisA collection of resources and tools for video analysis using deep learning and multi-modal learning techniques.767
drewnoakes/metadata-extractor-dotnetA .NET library for extracting metadata from various image, video, and audio file formats.953
rozumden/defmoA deep learning framework for deblurring and recovering the shape of fast-moving objects from blurred images171
tinghuiz/sfmlearnerA framework for unsupervised depth and ego-motion estimation from monocular videos using deep learning1,977
vision-cair/longvuAn artificial intelligence system designed to understand and describe long-form video content329
liuzhao1225/youdub-webuiA web-based video processing tool that uses AI to facilitate cultural and linguistic tasks such as transcription, translation, and audio synthesis.1,980
wangguanzhi/ladnA deep learning-based framework for facial makeup transfer and removal using adversarial disentangling networks182
adbedada/ts-rasterExtracts and analyzes time-series characteristics from raster data using Python.4