Awesome Infra for AI › AI Safety & Guardrails

Justin0504/Aegis

⭐ 486 TypeScript repository created 2026-03-04

Aegis acts as a critical security layer for AI agents, operating as a pre-execution firewall that intercepts, classifies, and potentially blocks tool calls before they execute. This prevents unwanted or malicious actions such as data exfiltration, system commands, or SQL injection, which can arise from prompt interpretation or model hallucinations. It integrates seamlessly into existing agent frameworks (e.g., Langchain, Anthropic) with minimal code changes, often just one line. The platform offers a Compliance Cockpit for real-time monitoring, trace visualization, policy management, and cost tracking across all agents. Key features include a tamper-evident audit trail with hash chaining and optional signing, an agent registry for identity management and declared tool scope, and a comprehensive 'AEGIS Agent Threat Ontology' for classifying agent-specific threats. It also provides a detector plugin contract for custom security teams, an LLM egress proxy to route all LLM calls through its detection and audit chain, and universal SIEM sinks for integration with tools like Splunk and Datadog. Additionally, Aegis supports RFC 6962 transparency logs for verifiable audit records and budget guards for cost control, ensuring AI agent operations are secure, compliant, and cost-efficient.

https://github.com/Justin0504/Aegis

ai-agentsai-safetyanthropicaudit-traillangchainllm-observabilitymcppolicy-enginefirewallruntime-securitygovernancecompliancethreat-detection

Also in AI Safety & Guardrails

data-privacy-stack/presidio

Presidio is an open-source framework for detecting, redacting, masking, and anonymizing sensitive data (PII) across text, images, and structured data, leveraging NLP and customizable pipelines.

NVIDIA-NeMo/Guardrails

NVIDIA NeMo Guardrails is an open-source toolkit for adding programmable guardrails to LLM-based conversational applications, focusing on safety, security, and controlled dialog.

superagent-ai/superagent

Superagent is an open-source SDK providing safety features for AI applications, including prompt injection detection, PII redaction, repository scanning for threats, and red teaming capabilities fo...

Tencent/AI-Infra-Guard

AI-Infra-Guard is a full-stack AI red teaming platform providing comprehensive security analysis, vulnerability scanning, and jailbreak evaluation for AI ecosystems and LLMs.

microsoft/agent-governance-toolkit

AI Agent Governance Toolkit (AGT) provides policy enforcement, identity management, execution sandboxing, and reliability engineering to secure autonomous AI agents in production.

FailproofAI/failproofai

Observability and policy enforcement for AI agent harnesses, hooking twelve coding and chat harnesses to record every run and block dangerous tool calls before they execute.

protectai/llm-guard

LLM Guard is a comprehensive open-source security toolkit designed to fortify Large Language Model (LLM) interactions by providing robust sanitization, malicious content detection, data leakage pre...

lennney/stop-that-shit

Multi-platform hook and skill guard for AI coding agents that blocks unrequested work such as generated hashes, checksums and task-scope expansion at the harness hook boundary.