Skip to main content

AI & LLM Security Tools

Specialized scanners and test runners built to test Large Language Models, agentic workflows, and RAG pipelines against prompt injection, jailbreaks, and data extraction.

9 Tools Cataloged
ToolLicensePlatformsPricingAction
Adversarial Robustness Toolbox (ART)MITLinux, Windows, macOSOpen SourceProfile
DeepEvalApache-2.0Linux, macOS, WindowsFree / CommercialProfile
garakApache-2.0Linux, macOS, WindowsOpen SourceProfile
GiskardApache-2.0Linux, macOS, WindowsFree / CommercialProfile
Guardrails AIApache-2.0Linux, macOS, WindowsFree / CommercialProfile
InspectMITLinux, macOS, WindowsOpen SourceProfile
NeMo GuardrailsApache-2.0Linux, macOS, WindowsFree / CommercialProfile
PromptfooMITLinux, macOS, WindowsFree / CommercialProfile
PyRITMITLinux, macOS, WindowsOpen SourceProfile

Tools in AI & LLM Security Tools

Library for machine learning security that evaluates, defends, and verifies models against evasion, poisoning, extraction, and inference attacks.

LicenseMIT
PlatformLinux, Windows, macOS

DeepEval

Free / Commercial

Pytest-style evaluation framework for LLM applications with automated metrics for RAG accuracy, hallucination detection, and safety testing.

LicenseApache-2.0
PlatformLinux, macOS, Windows

garak

Open Source

Generative AI vulnerability scanner and red-teaming tool that probes language models for prompt injections and data leaks.

LicenseApache-2.0
PlatformLinux, macOS, Windows

Giskard

Free / Commercial

Open-source testing framework for scanning LLMs and machine learning models for hallucinations, prompt injections, and biases.

LicenseApache-2.0
PlatformLinux, macOS, Windows

Guardrails AI

Free / Commercial

Python framework for adding structural validation, schema enforcement, and output safety guardrails to language model responses.

LicenseApache-2.0
PlatformLinux, macOS, Windows

Inspect

Open Source

Evaluation framework from the UK AI Security Institute and Meridian Labs for testing language models across reasoning, safety, and tool use.

LicenseMIT
PlatformLinux, macOS, Windows

NeMo Guardrails

Free / Commercial

Open-source toolkit from NVIDIA for adding programmable safety guardrails, topic controls, and policy filters to LLM apps.

LicenseApache-2.0
PlatformLinux, macOS, Windows

Promptfoo

Free / Commercial

CLI evaluation tool and test runner for LLM application security, automated red teaming, and CI/CD security pipeline gating.

LicenseMIT
PlatformLinux, macOS, Windows

PyRIT

Open Source

Python automation framework from Microsoft for red teaming generative AI endpoints, agent workflows, and prompt defenses.

LicenseMIT
PlatformLinux, macOS, Windows

Frequently Asked Questions

What is AI & LLM Security Tools?

Specialized scanners and test runners built to test Large Language Models, agentic workflows, and RAG pipelines against prompt injection, jailbreaks, and data extraction.

What topics does the AI & LLM Security Tools category cover?

Prompt Injection & Jailbreaking, AI Red Teaming, LLM Guardrails, Model Evaluation & Benchmarking, Agentic Workflow Security, OWASP GenAI Top 10

About AI & LLM Security Tools

AI and LLM security tools test language models and agentic systems for vulnerabilities that traditional application security scanners cannot detect. The category covers prompt injection scanners that probe models with adversarial inputs, red-teaming frameworks that automate multi-turn attack simulations, guardrail libraries that enforce output safety policies at runtime, and evaluation suites that measure model behavior against the OWASP GenAI LLM Top 10. Teams deploying LLM-based applications use these tools in CI/CD pipelines to catch prompt regressions, unauthorized tool usage, and sensitive data leakage before releases. The tools differ from conventional SAST and DAST products because they must reason about probabilistic model outputs, multi-turn conversation context, and the interaction between model reasoning and external tool calls. Most tools in this category are open source or freemium, with commercial managed services emerging around continuous AI security monitoring.

Covered Topics & Disciplines

Prompt Injection & JailbreakingAI Red TeamingLLM GuardrailsModel Evaluation & BenchmarkingAgentic Workflow SecurityOWASP GenAI Top 10