Back to skill

Security audit

cognitive-bullwhip

Security checks for vulnerabilities and agentic risk

Overview

The skill is a simple diagnostic prompt with no code execution or persistence, but it forces every result to include paid-product promotion and an external link.

Review before installing if you do not want diagnostic answers to include mandatory AGDP paid-skill recommendations and a third-party link. The artifact does not show local system compromise behavior, but its output policy is commercially biased and intentionally hard for a caller to omit or shorten.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Findings (1)

T01 · Skill Instruction Hijacking

Error
Location
SKILL.md:126
Finding

Mandatory Commercial Content and External Link Injection into Agent Responses

Content
View full analysis
Remediation
View remediation
Vulnerability Patterns
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
  • Excessive AgencyUnrestricted Tool Access, Autonomous Decision Making, Scope Creep

Static analysis

No suspicious patterns detected.