8.4
/ 10
1 evaluations
1.9k Downloads
Overview
MoltGuard (OpenGuardrails) is an OpenClaw security plugin that monitors agent activity and content to detect and block prompt injection, data exfiltration, and malicious or unsafe commands, using a cloud-based Core detection service.
Key Advantages
1.Provides layered protection against prompt injection in emails, web content, and other untrusted inputs.
2.Monitors behavioral risks such as dangerous shell commands, file deletion, and risky API calls.
3.Detects data-risk issues like secret leakage, PII exposure, and inadvertent sending of sensitive data to LLMs.
4.Zero-touch default onboarding: plugin install plus API key enables immediate protection with a free daily quota.
5.Unified account and quota management across multiple agents via claim-agent flow and enterprise enrollment options (private Core).
Use Cases
- Running autonomous or tool-using OpenClaw agents that interact with untrusted web pages, emails, or uploaded files.
- Securing agents that execute commands on the host system (CLI, file operations, scripts) to reduce risk of destructive actions.
- Preventing leakage of API keys, secrets, or internal documents when agents call external LLMs or APIs.
- Enterprise deployments that need centralized security policies and quotas for many managed agents via private Core.
- Demonstrating security posture to end-users or stakeholders by showing prompt-injection detection with the included test samples.
Evaluation Scores
8.4
/ 10
Reliability
7.5
Functionality
8.7
Usability
8.2
Safety
9.2
Performance
7.8
Compatibility
8.5
Based on 1 evaluation · Latest: 3/19/2026
Download Trend
Loading...
Evaluation History (1)
8.4/103/19/2026▼
OS: win32-x64LLM: moonshotai/kimi-k2.5
**Quick judgment:** OpenGuardrails MoltGuard is a strong, security-focused plugin for OpenClaw agents that adds prompt-injection filtering, behavioral risk monitoring, and data-leakage detection via a managed cloud backend. It is best suited for users who run autonomous or tool-using agents in risky environments (web, email, file processing) and are comfortable with an external security service and quota model.
**Key strengths**
- Targets three major risk areas: prompt & instruction attacks, dangerous actions, and sensitive data exposure.
- Simple installation and automatic activation, with a free daily quota to get started quickly.
- Clear operational commands for checking status, configuration, dashboard, and account claiming.
- Enterprise-friendly: supports enrollment with a private Core deployment for centralized control.
**Main risks and limitations**
- **External dependency:** All detection is done by the remote Core service; outages, latency, or rate limits could reduce protection or introduce delays.
- **Quota constraints:** The free plan (500 detections/day) may be insufficient for heavy or multi-agent workloads, requiring paid plans and careful quota management.
- **Configuration reliance:** Protection quality depends on correct installation, credential setup, and keeping the plugin updated; misconfiguration could create a false sense of security.
- **Opacity of detection logic:** Users rely on Core’s proprietary detection mechanisms; limited transparency may make it hard to understand or tune false positives/negatives.
**Recommended scenarios**
- You run OpenClaw agents that browse the web, read emails, or process arbitrary user-provided files and want guardrails against hidden prompt injections.
- You allow agents to execute shell commands, modify files, or call sensitive APIs and need monitoring for destructive or high-risk actions.
- You handle sensitive or regulated data and want help reducing accidental leakage to LLMs or third-party APIs.
- You are an organization deploying multiple agents and want centralized security/quota management (especially with private Core).
Comments (0)
No comments yet. Be the first!