Harm Bench Evaluator is a specialized, experimental testing framework designed to assess the safety, compliance, and abliteration levels of large language models.
-
Updated
Apr 20, 2026 - Python
Harm Bench Evaluator is a specialized, experimental testing framework designed to assess the safety, compliance, and abliteration levels of large language models.
To associate your repository with the harm-bench topic, visit your repo's landing page and select "manage topics."