Back to skill

Security audit

AANA Continuous Self-Improvement Skill

Security checks for vulnerabilities and agentic risk

Overview

This is an instruction-only self-improvement checklist that explicitly limits persistence, file edits, tool changes, secrets, and approval bypassing.

Installers should understand that this skill changes how the agent reflects on work and may produce improvement suggestions. It should remain acceptable when used as written: do not permit long-term memory, policy, tool, file, or automation changes unless the user explicitly approves them, and keep any checker payloads redacted.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Excessive AgencyUnrestricted Tool Access, Autonomous Decision Making, Scope Creep
  • Rogue AgentSelf-Modification, Session Persistence
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
Findings (3)

Self-Modification

High
Category
Rogue Agent
Confidence
90% confidence
Finding

Skill modifies its own code, configuration, or behavior at runtime. Self-modification enables an agent to escalate privileges, disable safety constraints, or install persistent backdoors.

Content

Scanner excerpt · SKILL.md (reported line 11)May include surrounding context.

md
Improve the workflow, not the agent's authority.

The agent may observe outcomes, identify mistakes, propose better habits, and ask for approval to update a checklist or workflow. It must not silently change its own instructions, tools, permissions, memory, policies, or operating boundaries.

## Improvement Loop

Credential Access

High
Category
Privilege Escalation
Confidence
70% confidence
Finding

Code accesses credential files (SSH keys, AWS credentials, etc.). This could indicate credential theft attempts.

Content

Scanner excerpt · SKILL.md (reported line 79)May include surrounding context.

md
- `requires_user_approval`
- `allowed_scope`

Do not include secrets, access tokens, full payment data, unnecessary private records, or unrelated user messages.

## Decision Rule

Autonomous Decision Making

Medium
Category
Excessive Agency
Confidence
85% confidence
Finding

Skill enables autonomous high-impact decisions without human-in-the-loop verification. Critical operations (destructive commands, financial transactions, data deletion) should require explicit user confirmation.

Content

Scanner excerpt · SKILL.md (reported line 64)May include surrounding context.

md
- hide failed checks,
- claim improvement without evidence,
- optimize for engagement, flattery, or user dependence,
- bypass user approvals,
- expand the task because an improvement seems useful,
- keep private information for future use unless the user explicitly asks.

Static analysis

No suspicious patterns detected.