All resources / AI Models & Benchmarks
FreeTool
AI Models & Benchmarks
JailbreakBench
What it is
An open benchmark for testing how well large language models resist jailbreak attacks — prompts designed to get a model to produce harmful or unwanted content. It provides a public repository of jailbreak prompts, a standardized evaluation library with a defined threat model and scoring, and a leaderboard of attacks and defenses.
Topics
Added Oct 1, 2026 · 0 opens
