BenchMark·Hub

LLM-ABAP-Code-Generation-Benchmark

B待确认Benchmark
代码代码生成开放mit

发布方:timkoehne

热度9.2▼ 0.1
下载量 · 30天
109
Hugging Face
GitHub Stars
代码仓库
论文被引
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

LLM Benchmark ABAP Code Generation Dataset This dataset is designed for benchmarking Large Language Models (LLMs) on ABAP code generation capabilities. It is based on the HumanEval benchmark, adapted for ABAP, and includes 16 additional ABAP-specific tasks that require interaction with database tables. Total tasks: 180 164 tasks adapted from HumanEval 16 ABAP-specific tasks Dataset Structure dataset.jsonl: Contains 180 examples. Each example has: id: Unique… See the full description on the dataset page: https://huggingface.co/datasets/timkoehne/LLM-ABAP-Code-Generation-Benchmark.