StructEval
B待确认Benchmark推理结构化输出开放mit
发布方:TIGER-Lab
热度22.4▼ 0.2
下载量 · 30天
795
Hugging Face
GitHub Stars
—
代码仓库
论文被引
35
Semantic Scholar
跑分模型 · 30天
—
Leaderboard results
简介
StructEval: A Benchmark for Structured Output Evaluation in LLMs StructEval is a benchmark dataset designed to evaluate the ability of large language models (LLMs) to generate and convert structured outputs across 18 different formats, and 44 types of tasks. It includes both renderable types (e.g., HTML, LaTeX, SVG) and non-renderable types (e.g., JSON, XML, TOML), supporting tasks such as format generation from natural language prompts and format-to-format conversion.… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/StructEval.