Back to skill

Security audit

polymarket-predictradar-whale-alert-skills

Security checks for vulnerabilities and agentic risk

Overview

The skill is mostly read-only and purpose-aligned, but its fallback paths can broaden results and show unverified wallets in a way that may mislead users relying on whale/smart-money signals.

Before installing, confirm you are comfortable with a read-only Polymarket analytics skill that may automatically loosen thresholds or include unverified wallets during data outages. Treat any Unclassified results as raw diagnostics, not verified whale activity.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Trigger AbuseOverly Broad Trigger, Shadow Command Trigger, Keyword Baiting Trigger
  • MCP Tool PoisoningHidden Instructions, Unicode Deception, Parameter Description Injection
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
Findings (5)

Vague Triggers

Medium
Category
Not specified by scanner
Confidence
97% confidence
Finding

The skill mixes fixed trigger words with open-ended auto-trigger examples, but it does not define clear precedence or boundaries for activation. This ambiguity increases the chance the skill will run in contexts where the user intended a generic trading discussion rather than a Polymarket whale query, expanding the blast radius of mistaken tool use.

Content

No source excerpt is available for this finding.

Vague Triggers

Medium
Category
Not specified by scanner
Confidence
93% confidence
Finding

The skill mixes fixed trigger words with open-ended auto-trigger examples, but it does not define clear precedence or boundaries for activation. This ambiguity increases the chance the skill will run in contexts where the user intended a generic trading discussion rather than a Polymarket whale query, expanding the blast radius of mistaken tool use.

Content

No source excerpt is available for this finding.

Whitespace Padding

Medium
Category
Prompt Injection
Confidence
70% confidence
Finding

Large whitespace padding was detected (a block of blank lines or a long run of spaces). This can push injected instructions below or to the right of the visible area so a human reviewer never sees them while the agent still reads them. Manual review of the hidden content is recommended.

Content

Scanner excerpt · SKILL.md (reported line 286)May include surrounding context.

md
## Error Handling

| Scenario                         | Resolution                                                                                                                            |
| -------------------------------- | ------------------------------------------------------------------------------------------------------------------------------------- |
| MCP query timeout                | Reduce time window to 12h and retry once; if still failing, report error and exit                                                     |
| No orders >= $5k in 24h          | Lower threshold to $2,000 and retry; note the adjusted threshold in output                                                            |

Intent-Code Divergence

Medium
Category
Not specified by scanner
Confidence
91% confidence
Finding

The instruction to lower the threshold from $5,000 to $2,000 when no qualifying orders are found changes the meaning of the skill from large-order monitoring to broader trade monitoring. This can produce materially misleading results for users who rely on the skill to identify whales, especially because the downgrade occurs automatically in an error-handling path rather than by explicit user request.

Content

No source excerpt is available for this finding.

Intent-Code Divergence

Medium
Category
Not specified by scanner
Confidence
96% confidence
Finding

The error-handling path explicitly says to skip profile filtering and show raw data as "Unclassified" when classification data is empty, which contradicts the skill’s core requirement to show only verified smart-money whale addresses. In practice this can cause the agent to present unverified wallets as whale activity, weakening trust boundaries and enabling misleading or unsafe output during backend degradation.

Content

No source excerpt is available for this finding.

Static analysis

No suspicious patterns detected.