BenchMark·Hub

StructEval

B待确认Benchmark
推理结构化输出开放mit

发布方:TIGER-Lab

热度22.4▼ 0.2
下载量 · 30天
795
Hugging Face
GitHub Stars
代码仓库
论文被引
35
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

StructEval: A Benchmark for Structured Output Evaluation in LLMs StructEval is a benchmark dataset designed to evaluate the ability of large language models (LLMs) to generate and convert structured outputs across 18 different formats, and 44 types of tasks. It includes both renderable types (e.g., HTML, LaTeX, SVG) and non-renderable types (e.g., JSON, XML, TOML), supporting tasks such as format generation from natural language prompts and format-to-format conversion.… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/StructEval.