Security audit
Arxiv Gamedevbench Evaluating Agentic Capabili
Security checks for vulnerabilities and agentic risk
Overview
This skill is a small, disclosed Node.js scaffold for an arXiv paper and does not show hidden access, persistence, network activity, or destructive behavior.
This appears safe to install as a simple generated research scaffold. Users should understand it is not a complete implementation of the paper, just a runnable placeholder that summarizes the paper and would need additional code for real experiments.
Vulnerability Patterns
- Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
- Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
- Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
- Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
- Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
- Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
- Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
- Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
- Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
- Excessive AgencyUnrestricted Tool Access, Autonomous Decision Making, Scope Creep
Static analysis
No suspicious patterns detected.
