Back to skill

Security audit

AI 写作风险评估

Security checks for vulnerabilities and agentic risk

Overview

This is a text-only AI-writing risk review skill with disclosed limits and no evidence of hidden access, persistence, or unsafe execution.

Installers should understand that this skill produces heuristic AI-writing risk assessments, not proof of authorship or misconduct. It is best suited for Chinese-language review workflows unless the user asks for another language, and any high-stakes use should rely on additional evidence such as drafts, sources, revision history, or human review.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Trigger AbuseOverly Broad Trigger, Shadow Command Trigger, Keyword Baiting Trigger
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
Findings (3)

Natural-Language Policy Violations

Medium
Category
Not specified by scanner
Confidence
95% confidence
Finding

The skill metadata and default prompt are written to force Chinese output regardless of the user's language preference. This can degrade usability, cause misunderstanding of sensitive evaluative results, and create consent/localization issues when the user expects another language or when downstream systems assume language alignment.

Content

No source excerpt is available for this finding.

Vague Triggers

Medium
Category
Not specified by scanner
Confidence
94% confidence
Finding

The default invocation prompt is generic and can be triggered on arbitrary text without clearly constraining when the skill should be used or what inputs are appropriate. In an agent environment, overly broad activation increases the chance of misrouting unrelated user requests into AI-authorship adjudication, producing unwarranted accusations, privacy-sensitive analysis, or unreliable conclusions outside the intended scope.

Content

No source excerpt is available for this finding.

Natural-Language Policy Violations

Medium
Category
Not specified by scanner
Confidence
94% confidence
Finding

This markdown template is entirely written as a Chinese-only reporting format and does not indicate that users may choose another language or locale. Under the stated policy, forcing a specific language without user opt-in is a natural-language policy violation unless the regional constraint is explicitly documented and justified.

Content

No source excerpt is available for this finding.

Static analysis

No suspicious patterns detected.