data-privacy-stack/presidio
Presidio is an open-source framework for detecting, redacting, masking, and anonymizing sensitive data (PII) across text, images, and structured data, leveraging NLP and customizable pipelines.
Awesome Infra for AI › AI Safety & Guardrails
Agent-Safe Pipeline implements the principle that an AI agent may propose an action but must not decide whether that action is authorized. The flow has four stages. Intent capture takes the agent proposal together with trusted context and freezes it into an immutable record, so what is evaluated is exactly what will run. That intent is submitted to an external policy service, Decionis, which returns one of three verdicts: allow, escalate or block. An escalation goes to a presence component that obtains verified human approval and returns the decision for re-evaluation rather than letting the approval bypass policy. Finally a SafeExecutor consumes a single-use grant bound to that specific intent and dispatches it through a sealed action registry that maps action names to trusted handlers and validates parameters. The executor never accepts an arbitrary callback from the agent, and the agent never holds the downstream privileged credentials or chooses which handler runs, which is what keeps the boundary meaningful. The repository is a runnable reference implementation and a set of packages rather than a hosted service, and it states plainly that its safety claims hold only while the documented trust boundary is preserved and that it does not replace provider-side identity, least privilege, network isolation or incident response. Packages cover the pipeline primitives, and examples walk from the smallest block flow upward. Supply chain hygiene is visible in the repository through continuous integration, code scanning, secret scanning and an OpenSSF Scorecard. It is aimed at teams giving agents the ability to call real APIs and needing a defensible record of who authorized each call.
https://github.com/decionis/agent-safe-pipeline
Presidio is an open-source framework for detecting, redacting, masking, and anonymizing sensitive data (PII) across text, images, and structured data, leveraging NLP and customizable pipelines.
NVIDIA NeMo Guardrails is an open-source toolkit for adding programmable guardrails to LLM-based conversational applications, focusing on safety, security, and controlled dialog.
Superagent is an open-source SDK providing safety features for AI applications, including prompt injection detection, PII redaction, repository scanning for threats, and red teaming capabilities fo...
AI-Infra-Guard is a full-stack AI red teaming platform providing comprehensive security analysis, vulnerability scanning, and jailbreak evaluation for AI ecosystems and LLMs.
AI Agent Governance Toolkit (AGT) provides policy enforcement, identity management, execution sandboxing, and reliability engineering to secure autonomous AI agents in production.
Observability and policy enforcement for AI agent harnesses, hooking twelve coding and chat harnesses to record every run and block dangerous tool calls before they execute.
LLM Guard is a comprehensive open-source security toolkit designed to fortify Large Language Model (LLM) interactions by providing robust sanitization, malicious content detection, data leakage pre...
Multi-platform hook and skill guard for AI coding agents that blocks unrequested work such as generated hashes, checksums and task-scope expansion at the harness hook boundary.