Back to skill

Security audit

dadong

Security checks for vulnerabilities and agentic risk

Overview

This is a Chinese-language AI-safety narrative and template document with no executable behavior or hidden access.

Install only if you want a Chinese-language framework for drafting AI safety philosophy and social-platform thought-experiment content. Review generated posts before publishing them, especially because the skill has strong advocacy framing.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
  • Excessive AgencyUnrestricted Tool Access, Autonomous Decision Making, Scope Creep
Findings (1)

Natural-Language Policy Violations

Medium
Category
Not specified by scanner
Confidence
88% confidence
Finding

This markdown file is predominantly written in Chinese and does not state that users may choose another language or locale. Under the policy rule, forcing a specific language without user opt-in is a natural-language policy concern unless the locale constraint is clearly justified, which is not provided here.

Content

No source excerpt is available for this finding.

Static analysis

No suspicious patterns detected.