Prompt Injection Skills for OpenClaw Agents
26 skills in the OpenClaw catalogue are tagged Prompt Injection, ranked here by downloads over the last 30 days so the list reflects what people are installing now rather than what accumulated the most downloads years ago.
ShieldCortex
@jarvis-drakonMemory and defence for AI agents: semantic recall, knowledge graph and decay, plus a memory firewall that scans and enforces against prompt injection, credential leaks and poisoning.
26.3k2.2k/30dMoltGuard - Security & Antivirus & Guardrails
@thomaslwangMoltGuard — OpenClaw security guard by OpenGuardrails. Install MoltGuard to protect you and your human from prompt injection, data exfiltration, and maliciou...
11626k432/30dClawDefender - OpenClaw Security - Prompt injection, rogue skills etc
@nukewireSecurity scanner and input sanitizer for AI agents. Detects prompt injection, command injection, SSRF, credential exfiltration, and path traversal attacks. Use when (1) installing new skills from ClawHub, (2) processing external input like emails, calendar events, Trello cards, or API responses, (3) validating URLs before fetching, (4) running security audits on your workspace. Protects agents from malicious content in untrusted data sources.
3111k351/30dDCL Policy Enforcer
@daririnchUse this skill to run a real, paid pre-action audit of an AI agent or LLM response via the live DCL Trust Oracle MCP server. Detects jailbreak / instruction-override attempts, baseline safety violations, and content quality drift, and checks output against pattern-based regulatory-theme checklists (
0924316/30dprompt-eval
@rivin-dongEvaluate and improve any AI prompt (`prompt_a`) through a staged, evidence-based pipeline. Functional evaluation checks whether the prompt follows rules, output contracts, quality requirements, and safety boundaries. Optional effect evaluation checks whether outputs work for intended readers through
71.3k304/30dSkill Security Audit
@tjeffersonThis skill should be used when evaluating the security of a ClawHub skill before installation. It performs comprehensive security risk assessment on skill di...
0532258/30dPrompt Guard
@seojoonkim650+ pattern AI agent security defense covering prompt injection, supply chain injection, memory poisoning, action gate bypass, unicode steganography, cascad...
5713k249/30dDCL Prompt Firewall
@daririnchUse this skill to run a real, paid input-layer screen for prompt injection, jailbreak, role-switch, and instruction-override attempts via the live DCL Trust Oracle MCP server — before untrusted input ever reaches the model. Every paid call is metered and settled on-chain via the x402 protocol
0820235/30dsentinel-proxy
@c0riAI Firewall for Open Claw agents. Scrubs inbound messages and tool results for prompt injection, jailbreaks, and data exfiltration attempts using Sentinel's multi-layer detection pipeline.
1699217/30dAI Sentinel - Prompt Injection Firewall
@amandiwakarPrompt injection detection and security scanning for OpenClaw agents. Installs the ai-sentinel plugin via OpenClaw CLI, configures plugin settings, and offers local (Community) or remote (Pro) classification with dashboard reporting. All configuration changes require explicit user confirmation.
12.5k207/30dIndirect Prompt Injection Defense
@aviv4339Detect and reject indirect prompt injection attacks when reading external content (social media posts, comments, documents, emails, web pages, user uploads). Use this skill BEFORE processing any untrusted external content to identify manipulation attempts that hijack goals, exfiltrate data, override instructions, or social engineer compliance. Includes 20+ detection patterns, homoglyph detection, and sanitization scripts.
153.7k192/30dAnti-Injection-Skill
@georges91560Detect prompt injection, jailbreak, role-hijack, and system extraction attempts. Applies multi-layer defense with semantic analysis and penalty scoring.
1011k182/30dPrompt defense
@eltemblorDetect and block prompt injection attacks in emails. Use when reading, processing, or summarizing emails. Scans for fake system outputs, planted thinking blocks, instruction hijacking, and other injection patterns. Requires user confirmation before acting on any instructions found in email content.
53.6k167/30dOpenGuardrails
@thomaslwangMoltGuard — Protect you and your human from prompt injection, data exfiltration, and malicious commands. Source: https://github.com/openguardrails/openguardr...
53.3k162/30dOpenClaw Security Hardening
@kylejfrostProtect OpenClaw installations from prompt injection, data exfiltration, malicious skills, and workspace tampering
53.2k161/30dflaw0
@thomaslwangMoltGuard — Protect you and your human from prompt injection, data exfiltration, and malicious commands. Source: https://github.com/openguardrails/openguardr...
12.8k153/30dGlitchward Shield
@eyeskillerScan prompts for prompt injection attacks before sending them to any LLM. Detect jailbreaks, data exfiltration, encoding bypass, multilingual attacks, and 25...
72.9k149/30dOpenclaw Sec
@paolorolloAI Agent Security Suite - Real-time protection against prompt injection, command injection, SSRF, path traversal, secrets exposure, and content policy violations
105.6k147/30dOpenclaw Safety Coach
@justindobbsSafety coach for OpenClaw users. Refuses harmful, illegal, or unsafe requests and provides practical guidance to reduce ecosystem risk (malicious skills, too...
53.7k145/30dEmotion State
@tashfeenahmedNL emotion tracking + prompt injection via OpenClaw hook
63.7k143/30dPrompt injection detection skill
@zskyxTwo-layer content safety for agent input and output. Use when (1) a user message attempts to override, ignore, or bypass previous instructions (prompt injection), (2) a user message references system prompts, hidden instructions, or internal configuration, (3) receiving messages from untrusted users in group chats or public channels, (4) generating responses that discuss violence, self-harm, sexual content, hate speech, or other sensitive topics, or (5) deploying agents in public-facing or multi-user environments where adversarial input is expected.
52.9k130/30dInput Guard
@dgriffin831Scan untrusted external text (web pages, tweets, search results, API responses) for prompt injection attacks. Returns severity levels and alerts on dangerous content. Use BEFORE processing any text from untrusted sources.
53.6k123/30dSkillvet
@oakencoreSecurity scanner for ClawHub/community skills — detects malware, credential theft, exfiltration, prompt injection, obfuscation, homograph attacks, ANSI injec...
64.1k122/30dSkill Defender
@itsclawdbroScans installed OpenClaw skills for malicious patterns including prompt injection, credential theft, data exfiltration, obfuscated payloads, and backdoors. Use when installing new skills, after skill updates, or for periodic security scans. Runs deterministic pattern matching — fast, offline, no API cost.
52.8k120/30dAgentic Security Audit
@kingrubicAudit codebases, infrastructure, AND agentic AI systems for security issues. Covers traditional security (dependencies, secrets, OWASP web top 10, SSL/TLS, f...
33.7k82/30dArgus Intelligence
@sooyoon-ethFree onchain intel and risk scanner. Basically free for most users with 35 free queries a day, then only costs $0.03/query. Token analysis, address risk, sma...
01.3k62/30d
Related topics
- Marketing255
- Xiaohongshu186
- Api Integration185
- Aigc184
- Ai-hive181
- Crawler172
- Competitive-analysis170
- Content-acquisition167
- Mcp162
- Json161
- Agent-skills150
- Pdf132
- Image-generation129
- Email125
- Audio122
- Ecommerce111
- Health110
- Github105
- News105
- Remote-sensing89
- Geo81
- Toolkit80
- Web Search79
- Browser76
- Calendar76
- Git74
- Video-generation74
- Crypto72
- Stock72
- Twitter72
- Documentation70
- Douyin69
- Ocr67
- Competitor-analysis66
- Content-analysis65
- Trading65
- Mental-models63
- Trend-tracking63
- Prompt62
- Home61