aya_evaluation_suite
B待确认Benchmark写作对话对话写作开放apache-2.0
发布方:CohereLabs
热度31.0▼ 0.1
下载量 · 30天
3,038
Hugging Face
GitHub Stars
—
代码仓库
论文被引
225
Semantic Scholar
跑分模型 · 30天
—
Leaderboard results
简介
Dataset Summary Aya Evaluation Suite contains a total of 26,750 open-ended conversation-style prompts to evaluate multilingual open-ended generation quality.To strike a balance between language coverage and the quality that comes with human curation, we create an evaluation suite that includes: human-curated examples in 7 languages (tur, eng, yor, arb, zho, por, tel) → aya-human-annotated. machine-translations of handpicked examples into 101 languages → dolly-machine-translated.… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/aya_evaluation_suite.