Skip to content
#

gaurdrail

Here are 7 public repositories matching this topic...

A fail-closed runtime security layer for AI agents with prompt injection detection, tool-call protection, PII/secret sanitization, HITL controls, and output validation.

  • Updated Sep 8, 2026
  • Python

Fine-tuned Llama-Prompt-Guard-2-86M classifier that blocks prompt injection attacks before they reach an LLM. +33 recall points over the base model on out-of-distribution attacks, ~20ms inference, served via FastAPI.

  • Updated Jul 30, 2026
  • Python

Add this topic to your repo

To associate your repository with the gaurdrail topic, visit your repo's landing page and select "manage topics."

Learn more