Back to skill

Security audit

许君山人格视角

Security checks for vulnerabilities and agentic risk

Overview

This is a disclosed Chinese persona-writing skill with no executable code, data access, persistence, or hidden behavior.

Before installing, understand that this skill may shape ordinary reply-writing or judgment requests in the Xu Junshan persona because implicit invocation is enabled. It does not appear to run code, access private files, use credentials, or persist changes.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Trigger AbuseOverly Broad Trigger, Shadow Command Trigger, Keyword Baiting Trigger
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
Findings (2)

Vague Triggers

High
Category
Not specified by scanner
Confidence
96% confidence
Finding

The trigger phrases are broad enough to match ordinary conversational requests such as '帮我回消息' or '帮我做判断', which can cause the skill to activate in situations the user did not explicitly intend. In a persona skill that rewrites communication in someone's voice, unintended activation can lead to confusing impersonation-style responses, reduced user agency, and inappropriate application of the persona to unrelated tasks.

Content

No source excerpt is available for this finding.

Natural-Language Policy Violations

Medium
Category
Not specified by scanner
Confidence
88% confidence
Finding

The instruction states '默认以中文输出', which imposes a language default rather than offering a neutral language choice at runtime. Under the policy, forcing a specific language without user opt-in can be a locale/language policy violation unless clearly justified as region-specific, which is not established here.

Content

No source excerpt is available for this finding.

Static analysis

No suspicious patterns detected.