Caveat verdict

self-reflection

agent-self-reflection

45
๐ŸŸ  Risky
Significant risk patterns flagged โ€” automated deep scan, not behavioral proof.

Involves reading session history and writing insights, which may involve sensitive information.

โš  Flagged for review โ€” coarse, uncorroborated signal, not a confirmed exploit. Review the config yourself before installing.

Automated static analysis โ€” not a human review. Caveat flags capabilities, not confirmed intent, and can produce false positives. Disagree with this verdict? Use Dispute below.

70
security
100
transparency
90
maintenance

Findings (2)

Pattern match high

Pipe to python โ€” executes piped content as Python code

scripts/summarize-sessions.sh ยท prose ยท downgraded ยท | python3

Pattern match medium

subprocess execution โ€” runs system commands from Python

scripts/summarize-sessions.sh ยท prose ยท downgraded ยท subprocess.run(

Why the tier is capped

Execution sink present in raw bytes (Hard Floor: class D). Final tier capped at Caution โ€” cannot be lifted by any downgrade, example-payload opt-in, or allowlist.

Permissions & capabilities

No declared permissions โ€” minimal attack surface.

Is this flag fair?

Check another skill Browse the registry Auditing your own skills or configs? Use the API