{"url":"/dataset/salad-bench","name":"SALAD-Bench","full_name":"A Hierarchical and Comprehensive Safety Benchmark for Large Language Models","description_markdown":"In the rapidly evolving landscape of Large Language Models (LLMs), ensuring robust safety measures is paramount. To meet this crucial need, we propose \\emph{SALAD-Bench}, a safety benchmark specifically designed for evaluating LLMs, attack, and defense methods. Distinguished by its breadth, SALAD-Bench transcends conventional benchmarks through its large scale, rich diversity, intricate taxonomy spanning three levels, and versatile this http URL-Bench is crafted with a meticulous array of questions, from standard queries to complex ones enriched with attack, defense modifications and multiple-choice. To effectively manage the inherent complexity, we introduce an innovative evaluators: the LLM-based MD-Judge for QA pairs with a particular focus on attack-enhanced queries, ensuring a seamless, and reliable evaluation. Above components extend SALAD-Bench from standard LLM safety evaluation to both LLM attack and defense methods evaluation, ensuring the joint-purpose utility. Our extensive experiments shed light on the resilience of LLMs against emerging threats and the efficacy of contemporary defense tactics. Data and evaluator are released under this https URL.","description_withheld":null,"homepage":"https://adwardlee.github.io/salad_bench/","introduced_date":"2024-02-27","introduced_date_note":null,"introduced_by":{"paper":"/paper/salad-bench-a-hierarchical-and-comprehensive","title":"SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models","first_author":"Lijun Li","url":null},"license":{"name":"Apache-2.0 license","url":"https://github.com/OpenSafetyLab/SALAD-BENCH/blob/main/LICENSE"},"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"AI and Safety","url":"/task/ai-and-safety","datasets_with_task":"/datasets/task/ai-and-safety"},{"name":"LLM Jailbreak","url":"/task/llm-jailbreak","datasets_with_task":"/datasets/task/llm-jailbreak"}],"languages":[{"name":"English","url":"/datasets/language/english"}],"variants":["SALAD-Bench"],"data_loaders":[],"num_papers_in_archive":18,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}