BenchMark·Hub

LongICLBench

B待确认Benchmark
长上下文检索上下文开放mit

发布方:TIGER-Lab

热度23.9▲ 0.1
下载量 · 30天
161
Hugging Face
GitHub Stars
代码仓库
论文被引
389
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

This is the benchmark we adopt in our TMLR2025 paper Long-context LLMs Struggle with Long In-context Learning. Check out our leaderboard at https://huggingface.co/spaces/TIGER-Lab/LongICL-Leaderboard. Cite our work by @misc{li2024longcontext, title={Long-context LLMs Struggle with Long In-context Learning}, author={Tianle Li and Ge Zhang and Quy Duc Do and Xiang Yue and Wenhu Chen}, year={2024}, eprint={2404.02060}, archivePrefix={arXiv}… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/LongICLBench.