Back to skill

Security audit

Model Replacement Prompt Generator

Security checks for vulnerabilities and agentic risk

Overview

This skill is a simple prompt generator with no executable code, but it forces uploaded-person replacement prompts to use a Chinese character even when the user did not ask for that.

Review before installing if you do not want automatic demographic constraints in image prompts. The skill appears technically low-risk, but users should only use it when they specifically want this replacement behavior and should consider editing the instructions to make identity traits optional and user-directed.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Trigger AbuseOverly Broad Trigger, Shadow Command Trigger, Keyword Baiting Trigger
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
Findings (3)

Natural-Language Policy Violations

High
Category
Not specified by scanner
Confidence
98% confidence
Finding

The skill hard-codes a protected ethnicity/nationality attribute ('Character must be Chinese') without user opt-in or task necessity. This creates discriminatory behavior in model outputs and can systematically force sensitive-attribute inference and transformation from uploaded images, which is unsafe and policy-relevant even if it is not a traditional exploit primitive.

Content

No source excerpt is available for this finding.

Natural-Language Policy Violations

High
Category
Not specified by scanner
Confidence
98% confidence
Finding

Repeating the mandatory requirement that the replacement character 'must be Chinese' reinforces a non-optional sensitive-attribute constraint throughout the skill logic. In the context of image-based person replacement, this increases the chance of biased or exclusionary outputs and normalizes transformation of people into a specific ethnicity regardless of user need or consent.

Content

No source excerpt is available for this finding.

Vague Triggers

Medium
Category
Not specified by scanner
Confidence
90% confidence
Finding

The listed triggers include generic phrases like "swap person" and "replace model," which can overlap with ordinary image-editing requests and do not clearly distinguish this skill from other editing workflows. The file also does not provide exclusion conditions or negative examples to clarify when the skill should not be used.

Content

No source excerpt is available for this finding.

Static analysis

No suspicious patterns detected.