BenchMark·Hub

AlGhafa-Arabic-LLM-Benchmark-Native

B待确认Benchmark
语言/语义多语低资源开放

发布方:OALL

热度31.0▲ 0.1
下载量 · 30天
6,732
Hugging Face
GitHub Stars
代码仓库
论文被引
354
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

AlGhafa Arabic LLM Benchmark New fix: Normalized whitespace characters and ensured consistency across all datasets for improved data quality and compatibility. Multiple-choice evaluation benchmark for zero- and few-shot evaluation of Arabic LLMs, we adapt the following tasks: Belebele Ar MSA Bandarkar et al. (2023): 900 entries Belebele Ar Dialects Bandarkar et al. (2023): 5400 entries COPA Ar: 89 entries machine-translated from English COPA and verified by native Arabic… See the full description on the dataset page: https://huggingface.co/datasets/OALL/AlGhafa-Arabic-LLM-Benchmark-Native.