LLM-ABAP-Code-Generation-Benchmark
B待确认Benchmark代码代码生成开放mit
发布方:timkoehne
热度9.2▼ 0.1
下载量 · 30天
109
Hugging Face
GitHub Stars
—
代码仓库
论文被引
—
Semantic Scholar
跑分模型 · 30天
—
Leaderboard results
简介
LLM Benchmark ABAP Code Generation Dataset This dataset is designed for benchmarking Large Language Models (LLMs) on ABAP code generation capabilities. It is based on the HumanEval benchmark, adapted for ABAP, and includes 16 additional ABAP-specific tasks that require interaction with database tables. Total tasks: 180 164 tasks adapted from HumanEval 16 ABAP-specific tasks Dataset Structure dataset.jsonl: Contains 180 examples. Each example has: id: Unique… See the full description on the dataset page: https://huggingface.co/datasets/timkoehne/LLM-ABAP-Code-Generation-Benchmark.