BenchMark·Hub

zest

B待确认Benchmark
推理通用推理开放cc-by-4.0

发布方:allenai

热度21.3▼ 0.3
下载量 · 30天
350
Hugging Face
GitHub Stars
代码仓库
论文被引
99
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

ZEST tests whether NLP systems can perform unseen tasks in a zero-shot way, given a natural language description of the task. It is an instantiation of our proposed framework "learning from task descriptions". The tasks include classification, typed entity extraction and relationship extraction, and each task is paired with 20 different annotated (input, output) examples. ZEST's structure allows us to systematically test whether models can generalize in five different ways.