BenchMark·Hub

VisPhyBench-Data

B待确认Benchmark
多模态视觉语言开放mit

发布方:TIGER-Lab

热度7.1▼ 0.1
下载量 · 30天
37
Hugging Face
GitHub Stars
代码仓库
论文被引
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

VisPhyBench To evaluate how well models reconstruct appearance and reproduce physically plausible motion, we introduce VisPhyBench, a unified evaluation protocol comprising 209 scenes derived from 108 physical templates that assesses physical understanding through the lens of code-driven resimulation in both 2D and 3D scenes, integrating metrics from different aspects. Each scene is also annotated with a coarse difficulty label (easy/medium/hard). Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/VisPhyBench-Data.