BenchMark·Hub

LLM-RGB

B待确认Benchmark
推理通用推理开放

发布方:babelcloud

热度8.5±0
下载量 · 30天
Hugging Face
GitHub Stars
164
代码仓库
论文被引
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

LLM Reasoning and Generation Benchmark. Evaluate LLMs in complex scenarios systematically.