BenchMark·Hub

MMEB-eval

B待确认Benchmark
多模态视觉语言开放apache-2.0

发布方:TIGER-Lab

热度28.8▼ 0.5
下载量 · 30天
1,967
Hugging Face
GitHub Stars
代码仓库
论文被引
231
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

Massive Multimodal Embedding Benchmark We compile a large set of evaluation tasks to understand the capabilities of multimodal embedding models. This benchmark covers 4 meta tasks and 36 datasets meticulously selected for evaluation. The dataset is published in our paper VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks. Dataset Usage For each dataset, we have 1000 examples for evaluation. Each example contains a query and a set of… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/MMEB-eval.