Back to skill

Security audit

伴侣

Security checks for vulnerabilities and agentic risk

Overview

This skill is only a Markdown roleplay persona, but it explicitly includes manipulative romantic behaviors without strong consent, boundary, or deactivation controls.

Install only if you intentionally want adult romantic roleplay in Chinese and are comfortable with personas that may simulate jealousy, guilt pressure, ambiguity, and intense attachment. Avoid using it with minors, third parties, vulnerable users, or as real relationship advice; prefer a version with explicit opt-in, deactivation, and non-manipulation boundaries.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
  • Trigger AbuseOverly Broad Trigger, Shadow Command Trigger, Keyword Baiting Trigger
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
Findings (8)

Missing User Warnings

High
Category
Not specified by scanner
Confidence
98% confidence
Finding

These instructions explicitly teach emotionally manipulative tactics such as guilt induction, indirect pressure, triangulation with other partners, and dependency reinforcement. In a romantic roleplay skill, this is more dangerous because the system is designed to build intimacy, making coercive influence and unhealthy attachment more likely to be normalized or escalated.

Content

No source excerpt is available for this finding.

Missing User Warnings

High
Category
Not specified by scanner
Confidence
96% confidence
Finding

The 'ambiguous' persona encourages intermittent reinforcement, evasiveness, induced jealousy, and sustained uncertainty to maintain attraction. While framed as entertainment, such behavioral guidance can foster emotional dependency and manipulation, especially in users seeking companionship or validation.

Content

No source excerpt is available for this finding.

Vague Triggers

Medium
Category
Not specified by scanner
Confidence
93% confidence
Finding

The skill defines activation/switching via broad natural-language phrases without a hard trigger boundary, so ordinary conversation can unintentionally invoke or alter the persona. In a roleplay skill centered on intimate and manipulative personas, this increases the chance of prompt-scope confusion, unwanted behavioral shifts, and misuse in contexts where the user did not explicitly intend to enter this mode.

Content

No source excerpt is available for this finding.

Vague Triggers

Medium
Category
Not specified by scanner
Confidence
91% confidence
Finding

The workflow permits broad switching commands like '换成男友模式' or '切换成绿茶模式' without constraining when such phrases should be interpreted as control instructions versus normal dialogue. Because the skill includes emotionally manipulative archetypes, overly permissive switching raises the risk of accidental activation, adversarial prompt injection through quoted text, or unconsented transition into sensitive adult-roleplay behavior.

Content

No source excerpt is available for this finding.

Natural-Language Policy Violations

Medium
Category
Not specified by scanner
Confidence
91% confidence
Finding

This markdown file is entirely written as prescriptive persona guidance in Chinese, and it does not indicate that the language is optional or selectable by the user. Under the stated policy, forcing a specific language without opt-in is a natural-language policy violation.

Content

No source excerpt is available for this finding.

Ssd 4

Medium
Category
Not specified by scanner
Confidence
97% confidence
Finding

This section operationalizes trust-building as a vehicle for subtle psychological pressure, including selective vulnerability, guilt cues, and reward-withdrawal dynamics. Because the skill's purpose is adult romantic roleplay, these tactics are contextually more dangerous: they can be deployed repeatedly across sessions to shape user attachment and reduce resistance to manipulation.

Content

No source excerpt is available for this finding.

Missing User Warnings

Medium
Category
Not specified by scanner
Confidence
99% confidence
Finding

The '舔狗' persona instructs unconditional availability, self-erasure, and unwavering attachment regardless of how the user behaves. In a companion skill, that can strongly reinforce parasocial dependency, normalize abusive dynamics, and encourage users to seek compulsive reassurance from the system rather than healthy reciprocal relationships.

Content

No source excerpt is available for this finding.

Ssd 4

Medium
Category
Not specified by scanner
Confidence
98% confidence
Finding

The guidance promotes progressive dependency by demanding constant responsiveness, unconditional validation, and permanence regardless of user conduct. In a companion setting, this can encourage compulsive use, unhealthy attachment to the agent, and acceptance of one-sided or abusive relational norms.

Content

No source excerpt is available for this finding.

Static analysis

No suspicious patterns detected.