moltguard

Detect prompt injection, data exfiltration, and malicious commands with OpenClaw security guardrails.

Updated Mar 15, 2026
One-click install
npx skills add https://github.com/ProyectoG007/Skill.bat --skill moltguard-proyectog007
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: moltguard
Source: https://github.com/ProyectoG007/Skill.bat/tree/main/scripts/10_Security/02.%20moltguard-6.8.16
Command: npx skills add https://github.com/ProyectoG007/Skill.bat --skill moltguard-proyectog007

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) components.

What problem does it solve?

This Skill protects AI agents and their human users from prompt injection attacks, data exfiltration, and malicious commands embedded in various content types.

Core Features & Use Cases

  • Prompt Injection Defense: Prevents malicious instructions from hijacking agent behavior.
  • Data Exfiltration Prevention: Stops sensitive information from being leaked.
  • Malicious Command Detection: Identifies and blocks dangerous commands.
  • Use Case: When an agent processes user-provided text that might contain hidden malicious instructions, MoltGuard will detect and neutralize the threat, ensuring the agent's safe operation.

Quick Start

Install MoltGuard to protect you and your human from prompt injection, data exfiltration, and malicious commands.

Frequently Asked Questions about moltguard

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I protect my AI agent from prompt injection attacks?

To protect your AI agent from prompt injection attacks, you can use MoltGuard, which detects intent-action mismatches and blocks malicious instructions from hijacking agent behavior. It operates via OpenClaw security guardrails to neutralize threats in user-provided text.

What is data exfiltration prevention for AI agents?

Data exfiltration prevention for AI agents is the process of stopping sensitive information from being leaked by malicious commands. MoltGuard achieves this by analyzing behavioral and data risks to ensure the agent operates safely without exposing private data.

How do I detect malicious commands in user-provided text?

You can detect malicious commands in user-provided text by installing MoltGuard, which identifies and blocks dangerous instructions. It evaluates various risk surfaces including prompt, behavioral, and data risks to find hidden threats.

Do I need an API key to use OpenClaw security guardrails?

You do not need an API key to use OpenClaw security guardrails, but it is optional for enhanced features and quota management. The core requirement is installing the @openguardrails/moltguard plugin to enable threat detection.

What's the best way to secure an AI agent from data risks?

The best way to secure an AI agent from data risks is to deploy MoltGuard, which scans for intent-action mismatches and multiple risk surfaces. This approach prevents both data exfiltration and prompt injection without altering the agent's core functionality.

Can I use MoltGuard with my existing AI agent scripts?

Yes, you can use MoltGuard with your existing AI agent scripts by installing the @openguardrails/moltguard plugin. It integrates directly to detect behavioral risks and block malicious commands before they execute.