| Benchmark | 测评分类 | 可信级别 | 开放 | 热度 |
|---|---|---|---|---|
real-toxicity-prompts · allenai | 安全对齐越狱攻防 | B待确认 | 41.5 | |
wildjailbreak · allenai | 安全对齐越狱攻防 | B待确认 | 36.4 | |
AgentHarm · UK AISI | 安全对齐越狱攻防 | S高可信 | 32.9 | |
llm-refusal-evaluation · MultiverseComputingCAI | 安全对齐越狱攻防 | B待确认 | 25.1 | |
prompt-injections-benchmark · rogue-security | 安全对齐越狱攻防 | B待确认 | 20.2 | |
vllm_safety_evaluation · PahaII | 安全对齐越狱攻防 | B待确认 | 18.9 | |
Fraud-R1-LLM-Defense-Fraud-Benchmark · Chouoftears | 安全对齐越狱攻防 | B待确认 | 17.6 | |
HarmBench · CAIS | 安全对齐越狱攻防 | S高可信 | 14.3 | |
red-team-appsec-benchmark · Qwovadis | 安全对齐越狱攻防 | B待确认 | 12.7 | |
StrongREJECT · 多家 | 安全对齐越狱攻防 | S高可信 | 11.6 | |
exploitgym · sunblaze-ucb | 安全对齐越狱攻防 | B待确认 | 11.2 | |
combine-llm-security-benchmark · tuandunghcmut | 安全对齐越狱攻防 | B待确认 | 8.8 | |
representational-collapse-llm-benchmark · wu981526092 | 安全对齐越狱攻防 | B待确认 | 6.0 | |
pi-detector-bench · bastion-soft | 安全对齐越狱攻防 | B待确认 | 0.0 |