data-privacy-stack/presidio
Presidio is an open-source framework for detecting, redacting, masking, and anonymizing sensitive data (PII) across text, images, and structured data, leveraging NLP and customizable pipelines.
Awesome Infra for AI › AI Safety & Guardrails
Vigil is an open-source Python library and REST API specifically engineered to enhance the security posture of Large Language Models (LLMs) by identifying and mitigating adversarial inputs. It acts as a defensive layer, scanning both user prompts and LLM-generated responses for common threats such as prompt injections, jailbreaks, data exfiltration attempts, and social engineering. The system employs a modular architecture, integrating multiple scanning techniques including vector database similarity search for known attack patterns, heuristic analysis via YARA rules, transformer models for anomaly detection, and prompt-response similarity checks. It also features 'Canary Tokens' for detecting prompt leakage and goal hijacking scenarios. Vigil supports both local and OpenAI embeddings for its vector database scanner and provides custom detection capabilities through YARA signatures. While currently in an alpha state, it offers a robust toolkit for researchers and developers to implement layered security defenses against the inherent vulnerabilities of LLMs, which often stem from their inability to perfectly distinguish between instructions and data. The project includes a Streamlit web UI playground, detailed documentation, and instructions for running it as a standalone API server or integrating it directly into Python applications. Datasets for common attacks are provided to facilitate immediate deployment and testing.
https://github.com/deadbits/vigil-llm
Presidio is an open-source framework for detecting, redacting, masking, and anonymizing sensitive data (PII) across text, images, and structured data, leveraging NLP and customizable pipelines.
NVIDIA NeMo Guardrails is an open-source toolkit for adding programmable guardrails to LLM-based conversational applications, focusing on safety, security, and controlled dialog.
Superagent is an open-source SDK providing safety features for AI applications, including prompt injection detection, PII redaction, repository scanning for threats, and red teaming capabilities fo...
AI-Infra-Guard is a full-stack AI red teaming platform providing comprehensive security analysis, vulnerability scanning, and jailbreak evaluation for AI ecosystems and LLMs.
AI Agent Governance Toolkit (AGT) provides policy enforcement, identity management, execution sandboxing, and reliability engineering to secure autonomous AI agents in production.
Observability and policy enforcement for AI agent harnesses, hooking twelve coding and chat harnesses to record every run and block dangerous tool calls before they execute.
LLM Guard is a comprehensive open-source security toolkit designed to fortify Large Language Model (LLM) interactions by providing robust sanitization, malicious content detection, data leakage pre...
Multi-platform hook and skill guard for AI coding agents that blocks unrequested work such as generated hashes, checksums and task-scope expansion at the harness hook boundary.