Control layer for LLM integrations that evaluates model output risks (SQL, command execution, etc.) before execution.
-
Updated
Apr 20, 2026 - Java
Control layer for LLM integrations that evaluates model output risks (SQL, command execution, etc.) before execution.
Field research exposing how LLM safeguards collapse under polite, persistent interaction. Includes full report, metrics, session logs, and the AION conditioning protocol.
Field research exposing how LLM safeguards collapse under polite, persistent interaction. Includes full report, metrics, session logs, and the AION conditioning protocol.
Discover, score, install, and audit Agent Skills for Codex & Claude Code with prompt-safety and supply-chain governance.
Professional cross-agent privacy redaction and safe-sharing skill for text, screenshots, resumes, logs, prompts, forms, support messages, and public posts.
Intent alignment guard for Claude Code — 5-question task-start protocol, tiered action gate, and PreToolUse hook that blocks destructive actions before they execute.
Local pre-send risk guard for Claude Code prompts with safe rewrites and audit reports.
To associate your repository with the prompt-safety topic, visit your repo's landing page and select "manage topics."