Caveat verdict

auto-arena

88
๐ŸŸข Trusted
No high-risk patterns surfaced by the deep scan โ€” automated capability review, not behavioral proof.

An automated LLM evaluation framework comparing models using user-configured endpoints and API keys; package_install is for standard ML evaluation dependencies and all network activity is for the stated benchmarking purpose.

โš  Flagged for review โ€” coarse, uncorroborated signal, not a confirmed exploit. Review the config yourself before installing.

Automated static analysis โ€” not a human review. Caveat flags capabilities, not confirmed intent, and can produce false positives. Disagree with this verdict? Use Dispute below.

50
security
20
transparency
70
maintenance

Permission integrity

Code accesses API keys/tokens but declares no environment variables

credential_access

Installs packages at runtime โ€” transitive dependencies are not auditable

package_install

Findings (1)

Pattern match critical

Possible hardcoded credential

SKILL.md ยท code ยท api_key: "${OPENAI_API_KEY}

Permissions & capabilities

No declared permissions โ€” minimal attack surface.

package_installnetwork_incredential_access

Is this flag fair?

Check another skill Browse the registry Auditing your own skills or configs? Use the API