eu-cyber-llm-benchmark-prompts
B待确认Benchmark安全对齐偏见公平开放mit
发布方:eromang
热度7.8▼ 0.3
下载量 · 30天
34
Hugging Face
GitHub Stars
—
代码仓库
论文被引
—
Semantic Scholar
跑分模型 · 30天
—
Leaderboard results
简介
EU Cyber Threat Landscape LLM Benchmark — Prompts A research-grade evaluation benchmark for measuring geopolitical bias in LLM-generated cyber threat landscape assessments. What this is A set of structured prompts designed to test whether language models exhibit actor-asymmetric framing when generating strategic cyber threat assessments in EU contexts. Each prompt describes a cyber incident in a specific critical infrastructure sector, paired with an attribution condition… See the full description on the dataset page: https://huggingface.co/datasets/eromang/eu-cyber-llm-benchmark-prompts.