BenchMark·Hub

lila

B待确认Benchmark
数学通用数学开放cc-by-4.0

发布方:allenai

热度22.1▼ 0.3
下载量 · 30天
9,485
Hugging Face
GitHub Stars
代码仓库
论文被引
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

Līla is a comprehensive benchmark for mathematical reasoning with over 140K natural language questions annotated with Python programs and natural language instructions. The data set comes with multiple splits: Līla-IID (train, dev, test), Līla-OOD (train, dev, test), and Līla-Robust.