Caveat verdict
arxiv-agentic-verifier
The skill explicitly executes code provided to it and includes a security warning recommending a sandbox; this is an intentional but broad capability that poses genuine risk without confirmed sandboxing.
⚠ Flagged for review — coarse, uncorroborated signal, not a confirmed exploit. Review the config yourself before installing.
Automated static analysis — not a human review. Caveat flags capabilities, not confirmed intent, and can produce false positives. Disagree with this verdict? Use Dispute below.
Findings (3)
References child_process — can spawn system processes
index.js · prose · downgraded · child_process
Accesses sensitive environment variables
index.js · prose · downgraded · process.env.OPENAI_API_KEY
Popular HTTP library — network access
index.js · prose · downgraded · Got
Why the tier is capped
Execution sink present in raw bytes (Hard Floor: class D). Final tier capped at Caution — cannot be lifted by any downgrade, example-payload opt-in, or allowlist.
Permissions & capabilities
No declared permissions — minimal attack surface.
Is this flag fair?
Thanks — recorded.