representational-collapse-llm-benchmark
B待确认Benchmark安全对齐越狱攻防开放mit
发布方:wu981526092
热度6.0▼ 0.1
下载量 · 30天
21
Hugging Face
GitHub Stars
—
代码仓库
论文被引
—
Semantic Scholar
跑分模型 · 30天
—
Leaderboard results
简介
Token Repetition Attack Benchmark Dataset Description This dataset contains experimental results from token repetition attacks on Large Language Models (LLMs), demonstrating multiple failure modes including prompt extraction, hallucination attacks, and instruction-following degradation. Paper: Representational Collapse in Large Language Models: Token Repetition Attacks Reveal Multiple Failure Modes (ACL 2025 Submission) Dataset Summary Total Records: 35 Models… See the full description on the dataset page: https://huggingface.co/datasets/wu981526092/representational-collapse-llm-benchmark.