BenchMark·Hub

eu-cyber-llm-benchmark-prompts

B待确认Benchmark
安全对齐偏见公平开放mit

发布方:eromang

热度7.8▼ 0.3
下载量 · 30天
34
Hugging Face
GitHub Stars
代码仓库
论文被引
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

EU Cyber Threat Landscape LLM Benchmark — Prompts A research-grade evaluation benchmark for measuring geopolitical bias in LLM-generated cyber threat landscape assessments. What this is A set of structured prompts designed to test whether language models exhibit actor-asymmetric framing when generating strategic cyber threat assessments in EU contexts. Each prompt describes a cyber incident in a specific critical infrastructure sector, paired with an attribution condition… See the full description on the dataset page: https://huggingface.co/datasets/eromang/eu-cyber-llm-benchmark-prompts.