Adversarial Robustness Toolbox (ART)
Linux Foundation AI & Data
Python library implementing adversarial attacks and defenses against machine learning models, covering evasion, poisoning, extraction and inference.
AI Security
Test and defend models, prompts, agents and the infrastructure around them.
45 tools profiled
How it differs Tests and guards models, LLM applications and agents against prompt injection, jailbreaks and data leakage. Scanning an AI app's own code is still SAST or SCA.
Linux Foundation AI & Data
Python library implementing adversarial attacks and defenses against machine learning models, covering evasion, poisoning, extraction and inference.
SplxAI
Open source CLI that statically analyzes agentic AI codebases and produces a visual map of agents, their tools and the risks in that wiring.
Tencent
Open source scanner that fingerprints self hosted AI infrastructure components and matches them against known vulnerabilities, with an MCP server analysis mode.
Akto
API security platform that builds an API inventory from mirrored traffic and runs automated tests for authorization, injection and data exposure flaws.
Praetorian
Open source prompt injection testing tool released by Praetorian for probing LLM applications with adversarial inputs.
Cerbos
Decoupled authorization engine that evaluates YAML policies over a principal, resource and action and returns an allow or deny decision over an API.
Cisco
Open source project from Cisco's AI Defense group addressing security controls for AI agent and LLM application activity.
Confident AI
Open source Python framework that generates adversarial prompts against an LLM application and scores the responses for vulnerabilities such as bias, PII leakage and excessive agency.
Linux Foundation AI & Data
Python library implementing adversarial attacks and defenses against machine learning models, covering evasion, poisoning, extraction and inference.
SplxAI
Open source CLI that statically analyzes agentic AI codebases and produces a visual map of agents, their tools and the risks in that wiring.
Tencent
Open source scanner that fingerprints self hosted AI infrastructure components and matches them against known vulnerabilities, with an MCP server analysis mode.
Akto
API security platform that builds an API inventory from mirrored traffic and runs automated tests for authorization, injection and data exposure flaws.
Praetorian
Open source prompt injection testing tool released by Praetorian for probing LLM applications with adversarial inputs.
Cerbos
Decoupled authorization engine that evaluates YAML policies over a principal, resource and action and returns an allow or deny decision over an API.
Cisco
Open source project from Cisco's AI Defense group addressing security controls for AI agent and LLM application activity.
Confident AI
Open source Python framework that generates adversarial prompts against an LLM application and scores the responses for vulnerabilities such as bias, PII leakage and excessive agency.
CyberArk
Open source fuzzer that applies a catalog of published jailbreak and prompt injection techniques against local or hosted language model endpoints.
NVIDIA
Command line LLM vulnerability scanner that fires a library of attack probes at a model endpoint and scores the responses with matched detectors.
Giskard
Python testing library that scans models and LLM applications for security and quality failures, then turns findings into a reusable test suite.
Guardrails AI
Python framework that wraps LLM calls in composable validators and decides what to do when a prompt or a response fails one.
Protect AI
Python library of composable input and output scanners that sanitize prompts and validate model responses entirely within your own environment.
Mindgard
Runs a library of AI attacks against your deployed models through their normal interface and reports which ones succeeded.
Promptfoo
Config driven test and red team harness for LLM applications, running assertions and generated adversarial probes against prompts, models and agents.
Microsoft
Python framework from Microsoft for automating adversarial probing of generative AI systems, with composable attack, transformation and scoring parts.
CyberArk
Open source fuzzer that applies a catalog of published jailbreak and prompt injection techniques against local or hosted language model endpoints.
NVIDIA
Command line LLM vulnerability scanner that fires a library of attack probes at a model endpoint and scores the responses with matched detectors.
Giskard
Python testing library that scans models and LLM applications for security and quality failures, then turns findings into a reusable test suite.
Guardrails AI
Python framework that wraps LLM calls in composable validators and decides what to do when a prompt or a response fails one.
Protect AI
Python library of composable input and output scanners that sanitize prompts and validate model responses entirely within your own environment.
Mindgard
Runs a library of AI attacks against your deployed models through their normal interface and reports which ones succeeded.
Promptfoo
Config driven test and red team harness for LLM applications, running assertions and generated adversarial probes against prompts, models and agents.
Microsoft
Python framework from Microsoft for automating adversarial probing of generative AI systems, with composable attack, transformation and scoring parts.