Back to skill

Security audit

Spanish

Security checks for vulnerabilities and agentic risk

Overview

This is a small writing-style skill for natural casual Spanish, with no code, network access, persistence, or data access.

Install this if you want more natural, informal Spanish writing. For formal or region-specific Spanish, explicitly tell the agent the audience, country, and register you want, especially because the skill includes casual slang and profanity examples.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
  • Excessive AgencyUnrestricted Tool Access, Autonomous Decision Making, Scope Creep
Findings (2)

Natural-Language Policy Violations

Medium
Confidence
95% confidence
Finding
The skill hard-codes Spanish output and a casual style without indicating user preference, which can override user intent, accessibility needs, or task requirements. This is not a classic security exploit, but it is a genuine safety/quality issue because it can cause unwanted language/register coercion and inappropriate outputs in professional or sensitive contexts.

Natural-Language Policy Violations

Medium
Confidence
90% confidence
Finding
These instructions explicitly enforce casual Spanish as the norm, which can produce mismatched tone in business, legal, medical, academic, or respectful interpersonal situations. In context, the issue is limited to style control rather than direct system compromise, but it can still degrade safety and appropriateness by biasing outputs toward informality without consent.

Static analysis

No suspicious patterns detected.