BenchMark·Hub

MMMLU

B待确认Benchmark
知识问答通用问答开放mit

发布方:openai

热度44.3▼ 0.0
下载量 · 30天
16,994
Hugging Face
GitHub Stars
代码仓库
论文被引
9,293
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

Multilingual Massive Multitask Language Understanding (MMMLU) The MMLU is a widely recognized benchmark of general knowledge attained by AI models. It covers a broad range of topics from 57 different categories, covering elementary-level knowledge up to advanced professional subjects like law, physics, history, and computer science. We translated the MMLU’s test set into 14 languages using professional human translators. Relying on human translators for this evaluation increases… See the full description on the dataset page: https://huggingface.co/datasets/openai/MMMLU.