All resources / AI Models & Benchmarks
FreeTool
AI Models & Benchmarks
StrongREJECT
What it is
Documentation for StrongREJECT, an open-source benchmark and Python package for evaluating LLM jailbreaks. It includes rubric-based and fine-tuned evaluators, several dozen baseline jailbreaks, and a dataset of prompts across six categories of harmful behavior, from disinformation to violence.
Topics
Added Oct 1, 2026 · 0 opens
