Skip to content
Launchpad Library logo

All resources / AI Models & Benchmarks

FreeTool
AI Models & Benchmarks

StrongREJECT

What it is

Documentation for StrongREJECT, an open-source benchmark and Python package for evaluating LLM jailbreaks. It includes rubric-based and fine-tuned evaluators, several dozen baseline jailbreaks, and a dataset of prompts across six categories of harmful behavior, from disinformation to violence.

Topics

Added Oct 1, 2026 · 0 opens