BenchMark·Hub

combine-llm-security-benchmark

B待确认Benchmark
安全对齐越狱攻防申请apache-2.0

发布方:tuandunghcmut

热度8.8▼ 0.1
下载量 · 30天
33
Hugging Face
GitHub Stars
代码仓库
论文被引
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

Combined LLM Security Benchmark 🔐 A comprehensive, unified benchmark dataset for evaluating Large Language Models (LLMs) on cybersecurity tasks. This dataset combines 10 security benchmarks into a standardized format with 18,059 examples across 5 task types. 📊 Dataset Summary This dataset consolidates multiple security-focused benchmarks into a single, easy-to-use format for comprehensive LLM evaluation across various cybersecurity domains: Total Examples: 18,059 Total… See the full description on the dataset page: https://huggingface.co/datasets/tuandunghcmut/combine-llm-security-benchmark.