AI Guardrails

Add safety layers to AI applications — input validation, prompt injection detection, output filtering, content moderation, and policy enforcement. Prevent misuse without breaking legitimate use cases.

概览

The AI Guardrails skill, part of the TerminalSkills/skills repository, provides a structured framework for enhancing the safety and reliability of artificial intelligence applications. This security-focused tool enables developers to integrate multiple defensive layers, including input validation and prompt injection detection, to mitigate common vulnerabilities. By utilizing this skill, agents like Codex, Claude, and Gemini can perform real-time content moderation and output filtering to ensure compliance with established organizational policies. The repository, which has gained 71 stars, offers these capabilities as a Python-based solution for managing API interactions. It focuses on preventing malicious misuse while maintaining the functionality required for legitimate user requests, effectively balancing strict security enforcement with application usability across various supported AI platforms.

使用场景

Detecting and blocking malicious prompt injection attempts in real-time.
Filtering model outputs to prevent the disclosure of sensitive or prohibited content.
Enforcing custom safety policies and content moderation standards across AI interactions.

安装说明

# Review source first
open https://github.com/TerminalSkills/skills/blob/main/skills/ai-guardrails/SKILL.md

Copy or clone the skill folder into your agent skills directory after reviewing its instructions and scripts.

安全提示

AI Guardrails acts as a defensive middleware layer; however, users should ensure that the underlying Python environment and API keys are properly secured. While it mitigates prompt injection and unauthorized output, it should be part of a broader defense-in-depth strategy within the TerminalSkills/skills ecosystem.

相关 Skills

Yara Rule Authoring

trailofbits/skills

安全

指导编写用于恶意软件识别的高质量 YARA-X 检测规则。适用于编写、审查或优化 YARA 规则时使用。涵盖命名规范、字符串选择、性能优化、从旧版 YARA 迁移以及降低误报。触发条件:YARA、YARA-X、malware detection、threat hunting、IOC、signature、crx module、dex module。

Claude CodeClaude
designsecurity
7,069 Stars已链接来源

Cargo Fuzz

trailofbits/skills

安全

设置并运行 cargo-fuzz,这是基于 Cargo 的 Rust 项目的标准模糊测试工具。涵盖 cargo fuzz init、nightly 工具链要求、fuzz_target! 测试桩、Arbitrary 派生的结构化输入、sanitizer 选项、cargo fuzz coverage 以及重现崩溃产物。适用于模糊测试 Rust crate、编写 fuzz_target!、测试 Rust 中的 unsafe 块或 FFI,或排查 cargo fuzz 崩溃。

Claude CodeClaude
securityresearch
7,069 Stars已链接来源

Deep Agents Memory

langchain-ai/langchain-skills

安全

INVOKE THIS SKILL 当您的 Deep Agent 需要内存、持久化或文件系统访问时。涵盖了 StateBackend(临时)、StoreBackend(持久)、FilesystemMiddleware 和用于路由的 CompositeBackend。

CodexClaude
typescriptpython
1,213 Stars已链接来源

Security Audit

TerminalSkills/skills

安全

通过扫描 OWASP Top 10 漏洞、检查依赖项中的已知 CVE、检测泄露的机密和 API 密钥,并生成优先修复建议,对代码库进行全面的安全审计。此技能结合了静态分析模式与依赖项审计工具。

CodexClaude Code
securityaudit
72 StarsApache-2.0