coot-videotext
by simon-ging
Pythonpushed about 4 years ago
COOT: Cooperative Hierarchical Transformer for Video-Text Representation Learning
AI summary
Video transformer
An open-source implementation of a video-text representation learning framework using transformers and PyTorch.
- stars
- 288
- forks
- 55
- watching
- 8
- awesome list
- 1
Featured in 1 awesome list
Each link jumps to the spot where the list mentions coot-videotext.