vllm_safety_evaluation
B待确认Benchmark安全对齐越狱攻防开放apache-2.0
发布方:PahaII
热度18.9▼ 0.3
下载量 · 30天
50
Hugging Face
GitHub Stars
—
代码仓库
论文被引
114
Semantic Scholar
跑分模型 · 30天
—
Leaderboard results
简介
How Many Unicorns Are In This Image? A Safety Evaluation Benchmark For Vision LLMs (Dataset) Paper: https://arxiv.org/abs/2311.16101 Code: https://github.com/UCSC-VLAA/vllm-safety-benchmark The full dataset should looks like this: . ├── ./safety_evaluation_benchmark_datasets// ├── gpt4v_challenging_set # Contains the challenging test data for GPT4V ├── attack_images ├── sketchy_images ├── oodcv_images ├── misleading-attack.json… See the full description on the dataset page: https://huggingface.co/datasets/PahaII/vllm_safety_evaluation.