Deterministic safety solutions for probabilistic AI agents
-
Updated
Jul 21, 2026 - Python
Deterministic safety solutions for probabilistic AI agents
Agent Execution Partnership AEE is an open-source control plane that ensures every AI agent action is authorized before it runs, observable while it runs, and verifiable after it completes.
Introducing XSafeClaw: The Open-Source Agent Safety Platform from Fudan University
Ethicore Engine™ is an AI safety, ethics, and compliance platform. This repo consists of the open-source components of Ethicore Engine™ - Guardian SDK; designed to protect your AI applications from prompt injection, jailbreaks, role hijacking, system-prompt extraction, and 100+ additional threat categories through a multi-layer analysis pipeline
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses | 500+ Papers | Perception, Cognition, Planning, Interaction, Agentic System
Runtime safety for AI coding agents with real-time enforcement, system-event monitoring, and long-horizon provenance. Supports Claude Code, Codex, Antigravity, Copilot, Omnigent on native macOS and Linux.
Practices, protocols, and skills for AI-driven software development. Skills and safety hooks for Claude Code, Codex, OpenCode, Cursor, Antigravity, and any agent supporting the Agent Skills standard.
An open taxonomy and scoring framework for evaluating AI agent sandboxes: 7 defense layers, 7 threat categories, 3 evaluation dimensions, 27 "sandboxes" scored.
Fast local Rust scanner for AI-agent prompt injection, credential leaks, exfiltration, and risky tool calls
The open standard for runtime agent control — declarative hooks, policy enforcement, and observability across AI agent frameworks.
Trust nothing. Ship safely. — Skeptical-reading and prompt-injection defense skill for AI agents. Provenance tagging, red-flag patterns, refusal templates, and a read-only injection auditor. MIT.
Human-in-the-loop execution for LLM agents
OpenClaw-compatible MASL safety gate with public RAG packs for memory-aware AI agents
The open-source safety layer for AI agents — block unsafe tool calls, require approval, enforce budgets, audit, replay.
🛡️ A curated list of tools, frameworks, standards, and resources for AI agent governance, safety, and compliance
Guardrails service for AI agents. Default-deny tool call evaluation with LLM safety analysis, priority-ordered decision matrix, and human-in-the-loop escalations. Session recording, behavioral analysis, MCP proxy, secret redaction, and real-time audit.
Security scanner for AI agent tool definitions
Runtime safety net for LLM agents. Detects token spirals, kills doomed tasks early, tells you exactly why. Rust core, Python SDK. pip install state-harness
Deterministic execution authorization for AI agents and automation systems
Add a description, image, and links to the agent-safety topic page so that developers can more easily learn about it.
To associate your repository with the agent-safety topic, visit your repo's landing page and select "manage topics."