BenchMark·Hub

SWE-QA-Pro-Bench

B待确认Benchmark
代码仓库级代码开放mit

发布方:TIGER-Lab

热度18.6▼ 0.5
下载量 · 30天
649
Hugging Face
GitHub Stars
代码仓库
论文被引
6
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

SWE-QA-Pro Bench (A Repository-level QA Benchmark Built from Diverse Long-tail Repositories) 💻 GitHub | 📖 Paper | 🤗 SWE-QA-Pro 📢 News 🚀 [2026-5-19] The evaluation code is released on GitHub. 🔥 [2026-3-23] SWE-QA-Pro Bench is publicly released! The model and code will be released soon. Introduction SWE-QA-Pro Bench is a repository-level question answering dataset designed to evaluate whether models can perform grounded, agentic reasoning… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/SWE-QA-Pro-Bench.