verifier.org

SPLX

publishes red-team findings against named frontier models, and open-sourced the agent-mapping half of its product.

not published#teams-wanting-to-map-an-

the most useful free artefact of any commercial vendor here, alongside published attacks on models people actually use.

splx publishes named red-team research against specific frontier models — gpt-5, claude opus 4.1 and grok 4 — plus an agentic red-teaming whitepaper. testing models people actually deploy, and saying what was found, is more useful than a vendor's own bypass-rate statistics.

agentic radar, its agent and tool-graph mapping tool, is open source on github and separable from the paid platform. for anyone trying to understand what their agent can reach before buying anything, that's a genuinely free starting point — though we did not fetch its licence file, so confirm the terms before depending on it.

the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.

pricing
not published
website
splx.ai
our verdict

the most useful free artefact of any commercial vendor here, alongside published attacks on models people actually use.

we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.

more ai red-teaming tools

NeMo Guardrails

nvidia's toolkit for constraining llm behaviour, used alongside testing

#guardrails#open-source

Giskard

open-source scanner for llm vulnerabilities and quality regressions

#open-source#scanner

ModelScan

scans model files for unsafe serialisation before you load them

#supply-chain#open-source

HiddenLayer

we checked thisnot published

fifty disclosed cves and thirty patents, across the broadest claimed scope here.

#enterprises-wanting-mode