#red-team
-
LLM Security Benchmarks Compared: What Each Measures
LLM security benchmarks compared: what HarmBench, JailbreakBench, AgentDojo, AgentHarm and CyberSecEval each measure, and where every one of them stops.
-
Best AI Security Testing Tools 2026: Scanners and Red Teams
A practitioner's comparison of the best AI security testing tools in 2026: open-source scanners, commercial red-teaming platforms, and how to choose.
-
Robust Intelligence (Now Cisco AI Defense): Platform Review
A review of Robust Intelligence, now part of Cisco AI Defense: algorithmic red teaming, model file scanning, and runtime protection of AI applications.
-
How to Evaluate AI Security Tools Without Getting Fooled
AI security tool demos are built for best-case scenarios. The evaluation dimensions, protocol, and vendor questions that expose how a tool really performs.
-
PyRIT Review: Microsoft's AI Red Teaming Framework
A review of PyRIT, Microsoft's open-source AI red teaming framework: its target, orchestrator, converter, scorer and memory design, and multi-turn attacks.
-
Garak LLM Scanner Review: Research Tool or CI Gate?
A review of garak, NVIDIA's open-source LLM vulnerability scanner: plugin architecture, backend coverage, report quality, and the CI-gating pattern.