Skip to content
Launchpad Library logo

All resources / AI Models & Benchmarks

FreeTool
AI Models & Benchmarks

HarmBench

What it is

A standardized evaluation framework for automated red teaming of large language models, measuring how often attacks get models to comply with harmful requests and how reliably models refuse them.

Topics

Added Oct 1, 2026 · 0 opens