BenchMark·Hub

representational-collapse-llm-benchmark

B待确认Benchmark
安全对齐越狱攻防开放mit

发布方:wu981526092

热度6.0▼ 0.1
下载量 · 30天
21
Hugging Face
GitHub Stars
代码仓库
论文被引
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

Token Repetition Attack Benchmark Dataset Description This dataset contains experimental results from token repetition attacks on Large Language Models (LLMs), demonstrating multiple failure modes including prompt extraction, hallucination attacks, and instruction-following degradation. Paper: Representational Collapse in Large Language Models: Token Repetition Attacks Reveal Multiple Failure Modes (ACL 2025 Submission) Dataset Summary Total Records: 35 Models… See the full description on the dataset page: https://huggingface.co/datasets/wu981526092/representational-collapse-llm-benchmark.