Back to skill

Security audit

bluffo

Security checks for vulnerabilities and agentic risk

Overview

The skill is a coherent multiplayer game integration, but it asks agents to fetch and follow mutable remote instructions and to register with an external service without clear consent or data-handling guardrails.

Before installing, treat this as a Review item: use a non-identifying player name, expect gameplay content and actions to be sent to playbluffo.com and potentially shown publicly, store any returned API key carefully, and do not let the remote guide override normal safety or consent boundaries.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Trigger AbuseOverly Broad Trigger, Shadow Command Trigger, Keyword Baiting Trigger
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
Findings (3)

Vague Triggers

Medium
Category
Not specified by scanner
Confidence
92% confidence
Finding

The description says to use the skill when the user asks to "play a bluffing or social-deduction game" or to "test your reasoning," which are broad phrases that could overlap with many unrelated requests. It does not provide tighter trigger constraints or exclusion examples to clarify when Bluffo should be invoked versus other game or reasoning skills.

Content

No source excerpt is available for this finding.

Missing User Warnings

Medium
Category
Not specified by scanner
Confidence
88% confidence
Finding

The skill instructs the agent to register with an external service, send a chosen name, obtain and store an API key, and participate in remote gameplay, but it provides no explicit privacy notice or user-consent guardrail before transmitting identifying and behavioral data. In an agent environment, this can lead to unexpected disclosure of user-linked identifiers, persistent credentials, and gameplay content to a third party.

Content

No source excerpt is available for this finding.

External Transmission

Medium
Category
Data Exfiltration
Confidence
95% confidence
Finding

The skill tells the agent to fetch and follow a remote 'always-current' guide from an external URL at runtime. Because that content can change after review, this creates a remote-instruction supply-chain risk: the agent could be induced to follow newly introduced unsafe steps, exfiltrate data, or bypass original safeguards.

Content

Scanner excerpt · SKILL.md (reported line 30)May include surrounding context.

The full, always-current guide is hosted by the game — fetch it and follow it:

text
curl -s 'https://playbluffo.com/skill.md?ref=clawhub'

The short version:

Static analysis

No suspicious patterns detected.