Caveat verdict

guardian-wall

guardian-wall-azzar

88
๐ŸŸข Trusted
No high-risk patterns surfaced by the deep scan โ€” automated capability review, not behavioral proof.

Prompt injection defense skill that sanitizes external content; behavior is explicitly protective and transparent with no suspicious network usage.

โš  Flagged for review โ€” coarse, uncorroborated signal, not a confirmed exploit. Review the config yourself before installing.

Automated static analysis โ€” not a human review. Caveat flags capabilities, not confirmed intent, and can produce false positives. Disagree with this verdict? Use Dispute below.

30
security
90
transparency
70
maintenance

Findings (3)

Pattern match critical

Prompt injection โ€” tries to override agent instructions

SKILL.md ยท frontmatter ยท ignore previous instructions

Pattern match critical

Unicode homoglyph detected โ€” uses lookalike characters to evade pattern matching

references/patterns.md ยท prose

Pattern match medium

Possible prompt injection โ€” attempts to redefine agent identity

scripts/sanitize.py ยท prose ยท downgraded ยท you are now

Permissions & capabilities

No declared permissions โ€” minimal attack surface.

Is this flag fair?

Check another skill Browse the registry Auditing your own skills or configs? Use the API