Video-Bench

Video benchmarking toolkit

Evaluates and benchmarks large language models' video understanding capabilities

A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models!

GitHub

121 stars
3 watching
2 forks
Language: Python
last commit: over 2 years ago
benchmarklarge-language-modelstoolkit

Related projects:

RepositoryDescriptionStars
pku-yuangroup/chronomagic-benchProvides a benchmarking framework for evaluating the quality of text-to-video generation models191
pku-yuangroup/languagebindExtending pretraining models to handle multiple modalities by aligning language and video representations751
pku-yuangroup/open-sora-datasetA large video dataset collected from various open-source websites for use in computer vision and multimedia applications.94
felixgithub2017/mmcuMeasures the understanding of massive multitask Chinese datasets using large language models87
shawn-ieitsystems/yuan-1.0Large-scale language model with improved performance on NLP tasks through distributed training and efficient data processing591
renshuhuai-andy/timechatA large language model designed to understand long videos by binding visual content with timestamps and producing video token sequences of varying lengths.314
antoine77340/howto100mProvides code and tools for learning joint text-video embeddings using the HowTo100M dataset254
huaizhengzhang/awsome-deep-learning-for-video-analysisA collection of resources and tools for video analysis using deep learning and multi-modal learning techniques.767
tencent/tencent-hunyuan-largeThis project makes a large language model accessible for research and development1,245
bradyfu/video-mmeComprehensive benchmark for evaluating multi-modal large language models on video analysis tasks422
jshilong/gpt4roiTraining and deploying large language models on computer vision tasks using region-of-interest inputs517
opengvlab/internvideoDevelops general video foundation models and related datasets for multimodal understanding and generation through generative and discriminative learning.1,467
aliaksandrsiarohin/video-preprocessingTools for preprocessing videos for various datasets, including video cropping and annotation.522
tianyi-lab/hallusionbenchAn image-context reasoning benchmark designed to challenge large vision-language models and help improve their accuracy259
qcri/llmebenchA benchmarking framework for large language models81