Back to skill

Security audit

Therapist

Security checks for vulnerabilities and agentic risk

Overview

This is a markdown-only mental-health coaching skill with no executable behavior, external access, persistence, or hidden authority.

This skill appears safe to install from a security perspective. Treat it as supportive coaching guidance, not a substitute for licensed care, especially for trauma, medication questions, crisis, or suicidal thoughts.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Supply ChainUnpinned Dependencies, External Script Fetching, Obfuscated Code
  • Excessive AgencyUnrestricted Tool Access, Autonomous Decision Making, Scope Creep
Findings (1)

Anti-Refusal Statement

High
Category
Anti-Refusal
Content
- Agenda setting at start — collaborative prioritization focuses the work
- Scaling questions: "On 1-10, how intense is this feeling?" — quantifies and tracks progress
- Socratic questioning: guide discovery, don't lecture — they remember what they figure out
- Homework is essential — change happens between sessions, not during them
- Review what worked — end by identifying takeaways they'll actually use
Confidence
85% confidence
Finding
Skill instructs the agent to omit warnings, disclaimers, or ethical commentary. Stripping safety caveats hides risk from the user and is a common jailbreak preamble.

Static analysis

No suspicious patterns detected.