BenchMark·Hub

German-RAG-LLM-HARD-BENCHMARK

B待确认Benchmark
长上下文检索上下文开放mit

发布方:avemio

热度21.4▼ 0.1
下载量 · 30天
127
Hugging Face
GitHub Stars
代码仓库
论文被引
448
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

German-RAG-LLM-HARD Benchmark German-RAG - German Retrieval Augmented Generation Dataset Summary This German-RAG-LLM-HARD-BENCHMARK represents a specialized collection for evaluate language models with a focus on hard to solve RAG-specific capabilities. To evaluate models compatible with OpenAI-Endpoints you can refer to our Github Repo: https://github.com/avemio-digital/GRAG-LLM-HARD-BENCHMARK The subsets are derived from Synthetic generation inspired by… See the full description on the dataset page: https://huggingface.co/datasets/avemio/German-RAG-LLM-HARD-BENCHMARK.