Back to skill

Security audit

Karpathy编程四大原则

Security checks for vulnerabilities and agentic risk

Overview

This is a Markdown-only coding-principles skill with broad Chinese triggers, but it does not install code, access sensitive data, persist, or execute commands.

Installers should expect this skill to influence ordinary coding conversations, especially Chinese prompts about writing code, fixing bugs, or adding features. It appears safe from a security perspective, but users who want opt-in-only behavior may prefer narrower trigger wording.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Trigger AbuseOverly Broad Trigger, Shadow Command Trigger, Keyword Baiting Trigger
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
Findings (3)

Vague Triggers

High
Category
Not specified by scanner
Confidence
97% confidence
Finding

The '禁用触发' section uses very broad phrases like '帮我写个...', '修复这个', and '添加功能', which are common in normal software conversations. This can cause the skill to over-trigger on routine requests and override user intent by forcing a prescribed workflow, creating prompt-scope interference and unpredictable agent behavior.

Content

No source excerpt is available for this finding.

Natural-Language Policy Violations

Medium
Category
Not specified by scanner
Confidence
83% confidence
Finding

The skill metadata and content are written entirely in Chinese and present Chinese trigger phrases as the activation interface, without offering a language choice or documenting that the skill is intentionally Chinese-only. This can violate language/locale policy if the skill implicitly forces a specific language on users.

Content

No source excerpt is available for this finding.

Vague Triggers

Medium
Category
Not specified by scanner
Confidence
93% confidence
Finding

The listed trigger phrase "编程法则" is generic, and the document later broadens activation to common user requests. This creates a risk of unintended invocation because ordinary coding conversations could match the trigger without clear boundaries.

Content

No source excerpt is available for this finding.

Static analysis

No suspicious patterns detected.