Caveat verdict

evalscope

skill-evalscope

88
๐ŸŸข Trusted
No high-risk patterns surfaced by the deep scan โ€” automated capability review, not behavioral proof.

Receives external input AND uses eval

LLM evaluation CLI using the published evalscope package; runs benchmarks against user-configured endpoints with user own keys.

Automated static analysis โ€” not a human review. Caveat flags capabilities, not confirmed intent, and can produce false positives. Disagree with this verdict? Use Dispute below.

45
security
30
transparency
70
maintenance

What it does

These are capability combinations: each listed behavior occurs in the skill, but Caveat detects co-occurrence โ€” it does not verify that one flows into another. Read the code to confirm a live chain.

Capability combination critical

Receives external input AND uses eval โ€” the remote code-injection pattern (data-flow not verified)

LLM01 ยท LLM05 ยท ASI01 ยท ASI05

Permission integrity

Installs packages at runtime โ€” transitive dependencies are not auditable

package_install

Findings (1)

Pattern match medium

HTTP request to bare IP address โ€” common in malicious payloads

perf-reference.md ยท prose ยท downgraded ยท http://127.0.0.1

Permissions & capabilities

No declared permissions โ€” minimal attack surface.

package_installnetwork_indynamic_evalcredential_access
Check another skill Browse the registry Auditing your own skills or configs? Use the API