Saturday, May 16, 2026
HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal