VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs
-
Updated
Feb 3, 2026 - Python
VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs
[ICLR 2026] RIVER: A Real-Time Interaction Benchmark for Video LLMs
UG2: a Video Benchmark for Assessing the Impact of Image Restoration and Enhancement on Automatic Visual Recognition
A curated, mathematically rigorous survey and resource catalog for unsupervised, self-supervised, and multimodal foundation-model video summarization.
Open-source frame-level video quality benchmark tool using VMAF, PSNR, and FFmpeg (libvmaf).
To associate your repository with the video-benchmark topic, visit your repo's landing page and select "manage topics."