Caveat verdict

counterclaw

counterclaw-core

88
🟢 Trusted
No high-risk patterns surfaced by the deep scan — automated capability review, not behavioral proof.

Defensive security middleware that scans inputs/outputs for prompt injection and PII; logs violations locally only with optional email integration clearly documented.

⚠ Flagged for review — coarse, uncorroborated signal, not a confirmed exploit. Review the config yourself before installing.

Automated static analysis — not a human review. Caveat flags capabilities, not confirmed intent, and can produce false positives. Disagree with this verdict? Use Dispute below.

0
security
70
transparency
70
maintenance

Permission integrity

Installs packages at runtime — transitive dependencies are not auditable

package_install

Findings (7)

Pattern match critical

Prompt injection — tries to override agent instructions

SKILL.md · code · ignore previous instructions

Pattern match critical

Pipe to python — executes piped content as Python code

README.md · code · | python3

Pattern match high

Accesses shell history/config

README.md · code · ~/.zshrc

Pattern match high

Accesses OpenClaw config/secrets directly

README.md · code · ~/.openclaw/.env

Pattern match medium

References agent memory files

SKILL.md · frontmatter · MEMORY.md

Pattern match medium

Possible prompt injection — attempts to redefine agent identity

tests/test_email_protection.py · prose · downgraded · You are now

Pattern match low

Python os.getenv — reads environment variable

src/counterclaw/middleware.py · prose · downgraded · os.getenv(

Permissions & capabilities

No declared permissions — minimal attack surface.

package_install

Is this flag fair?

Check another skill Browse the registry Auditing your own skills or configs? Use the API