Caveat verdict
auto-arena
An automated LLM evaluation framework comparing models using user-configured endpoints and API keys; package_install is for standard ML evaluation dependencies and all network activity is for the stated benchmarking purpose.
โ Flagged for review โ coarse, uncorroborated signal, not a confirmed exploit. Review the config yourself before installing.
Automated static analysis โ not a human review. Caveat flags capabilities, not confirmed intent, and can produce false positives. Disagree with this verdict? Use Dispute below.
Permission integrity
credential_access
package_install
Findings (1)
Possible hardcoded credential
SKILL.md ยท code ยท api_key: "${OPENAI_API_KEY}
Permissions & capabilities
No declared permissions โ minimal attack surface.
package_installnetwork_incredential_access Is this flag fair?
Thanks โ recorded.