garak checks whether a language model can be made to fail in ways you do not want. It probes for hallucination, data leakage, prompt injection, misinformation, toxicity, and jailbreaks, combining static, dynamic, and adaptive attacks. Think nmap or Metasploit, but for LLMs. It supports Hugging Face, OpenAI, Bedrock, NIM, litellm, and anything reachable over REST.
.jpg)
Testimonials
