# HarmBench

> Standardized evaluation framework for LLM robustness

- Category: [AI Security & Safety](https://ailandscape.org/category/ai-security-safety) › Offensive Testing & Red Teaming
- Homepage: https://harmbench.org
- Repository: https://github.com/centerforaisafety/HarmBench
- Tags: red-teaming, benchmarks, safety
- Added to the landscape: 2026-03-18

## Similar tools in Offensive Testing & Red Teaming

- [ART](https://ailandscape.org/tool/art): Adversarial Robustness Toolbox by IBM
- [CleverHans](https://ailandscape.org/tool/cleverhans): Library for benchmarking adversarial ML attacks
- [Foolbox](https://ailandscape.org/tool/foolbox): Fast adversarial attacks for deep neural networks
- [Garak](https://ailandscape.org/tool/garak): LLM vulnerability scanner and red-teaming framework
- [PyRIT](https://ailandscape.org/tool/pyrit): Python Risk Identification Toolkit for GenAI by Microsoft
- [TextAttack](https://ailandscape.org/tool/textattack): NLP adversarial attacks, augmentation, and training

---

Part of [AI Landscape](https://ailandscape.org), an open map of the AI ecosystem. Web page: https://ailandscape.org/tool/harmbench · Index for AI assistants: https://ailandscape.org/llms.txt
