BenchMark·Hub

VideoEval-Pro

B待确认Benchmark
多模态视频开放

发布方:TIGER-Lab

热度21.7▼ 0.1
下载量 · 30天
1,276
Hugging Face
GitHub Stars
代码仓库
论文被引
19
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

VideoEval-Pro VideoEval-Pro is a robust and realistic long video understanding benchmark containing open-ended, short-answer QA problems. The dataset is constructed by reformatting questions from four existing long video understanding MCQ benchmarks: Video-MME, MLVU, LVBench, and LongVideoBench into free-form questions. The paper can be found here. The evaluation code and scripts are available at: TIGER-AI-Lab/VideoEval-Pro Dataset Structure Each example in the… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/VideoEval-Pro.