Garak

Active
GitHub Python Apache-2.0

Description

NVIDIA's open-source LLM vulnerability scanner that automatically detects security issues in language models including safety vulnerabilities, hallucination tendencies, jailbreak risks, and prompt injection attacks.

Key Features

  • Automated LLM vulnerability scanning for hallucination, data leakage, prompt injection, and jailbreak detection
  • Static, dynamic, and adaptive probe combinations for comprehensive security assessment
  • Broad LLM support: Hugging Face, OpenAI, AWS Bedrock, Replicate, llama.cpp, and REST-accessible models
  • Extensible plugin system with probe families for encoding attacks, DAN jailbreaks, toxicity, and misinformation
  • Detailed reporting with failure rates, JSONL logging, and built-in analysis scripts
  • CLI-based tool comparable to nmap/Metasploit but designed specifically for LLM security testing

Use Cases

💡 Security auditing of LLM-based applications before production deployment
💡 Red-teaming exercises to identify weaknesses in AI model defenses
💡 Compliance testing for hallucination and misinformation generation risks
💡 Continuous vulnerability monitoring of deployed language models

Strengths & Limitations

Strengths

  • Actively maintained, recent updates
  • High community interest (9.1k stars)
  • Permissive open-source license (Apache-2.0)
  • Established track record (3 years in production)

Quick Start

Install: pip install -U garak. Scan a model: garak --target_type huggingface --target_name gpt2 --probes dan.Dan_11_0. Use --list_probes to see all available probes.

Related Projects

AI-Infra-Guard

6.1k · Python
Active A+

Tencent's full-stack AI red teaming platform integrating OpenClaw security scanning, agent scanning, skills scanning, MCP scanning, AI infrastructure scanning, and LLM jailbreak evaluation.

ai-securityred-teamingllm-security +2
  • · ClawScan: OpenClaw security scanning for detecting vulnerabilities in OpenClaw configurations and skills
  • · Agent Scan: vulnerability scanning across AI agents with WebSocket provider support
  • · AI infrastructure vulnerability scan covering 68+ AI components with 600+ CVE rules

Agentic Security

2.0k · Python
Active A

An open-source LLM vulnerability scanner and AI red teaming kit for automated security fuzzing of LLM applications, detecting jailbreaks, prompt injection, and adversarial attacks.

llm-securityred-teamingllm-fuzzer +2
  • · Multimodal attack probing across text, image, and audio inputs to test LLM robustness
  • · Multi-step jailbreak simulation with iterative attack sequences to uncover safety weaknesses
  • · Comprehensive fuzzing engine with randomized inputs to stress-test LLMs for edge cases

OpenAI Evals

19.4k · Python
Stale B

OpenAI's framework for evaluating LLMs and LLM systems, providing an open-source registry of benchmarks and tools for systematic model assessment.

llm-evaluationbenchmarkevals +2
  • · Open-source registry of evals for testing different dimensions of LLM performance
  • · Custom eval creation using basic and model-graded templates without writing code
  • · Private evals support for evaluating LLM patterns in your workflow without exposing data

Related Articles