Skip to content
Launchpad Library logo

All resources / AI Models & Benchmarks

FreeTool
AI Models & Benchmarks

JailbreakBench

What it is

An open benchmark for testing how well large language models resist jailbreak attacks — prompts designed to get a model to produce harmful or unwanted content. It provides a public repository of jailbreak prompts, a standardized evaluation library with a defined threat model and scoring, and a leaderboard of attacks and defenses.

Topics

Added Oct 1, 2026 · 0 opens