Back to skill

Security audit

DeepthinkLite

Security checks for vulnerabilities and agentic risk

Overview

DeepthinkLite is a straightforward local research workflow helper that creates disclosed Markdown and JSON artifacts without hidden network access, credential access, persistence, or destructive behavior.

Installers should expect this skill to create local research folders and may be prompted before web browsing or reading less obvious local paths. Use summary-only mode when working with sensitive or untrusted source material.

Vulnerability Patterns
  • Skill Instruction HijackingAlters the agent's session goals or safety constraints when the skill loads
  • Agent Memory PoisoningWrites attacker-controlled rules into memory that affect later sessions
  • Remote Payload Retrieval and ExecutionFetches external code whose behavior can change after review
  • Embedded Malicious CodeShips malicious scripts inside the skill and executes them locally
  • Unauthorized Access and Privilege EscalationObtains permissions beyond the task's legitimate needs
Vulnerability Patterns
  • Privilege EscalationExcessive Permissions, Sudo/Root Execution, Credential Access
  • Rogue AgentSelf-Modification, Session Persistence
  • MCP Least PrivilegeUnderdeclared Capability, Wildcard Permission, Missing Permission Declaration
  • Prompt InjectionInstruction Override, Hidden Instructions, Exfiltration Commands
  • Data ExfiltrationExternal Transmission, Env Variable Harvesting, File System Enumeration
Findings (3)

Self-Modification

High
Category
Rogue Agent
Confidence
85% confidence
Finding

Skill modifies its own code, configuration, or behavior at runtime. Self-modification enables an agent to escalate privileges, disable safety constraints, or install persistent backdoors.

Content

Scanner excerpt · SKILL.md (reported line 121)May include surrounding context.

md
## Scripts

- `deepthinklite query ...` creates the run directory + boilerplate.
- Safe to rerun: it will not overwrite existing files.

Credential Access

High
Category
Privilege Escalation
Confidence
70% confidence
Finding

Code accesses credential files (SSH keys, AWS credentials, etc.). This could indicate credential theft attempts.

Content

Scanner excerpt · scripts/deepthinklite.py (reported line 66)May include surrounding context.

python
"Safe handling rules:\n"
            "- Prefer summaries with citations.\n"
            "- If you include raw text, wrap it in a clearly delimited UNTRUSTED block.\n"
            "- Never access secrets/credentials. Ask permission before web browsing or reading non-obvious local paths.\n\n"
            "## Restatement\n\n"
            "(Restate the ask in 1–2 lines.)\n\n"
            "## Success criteria\n\n"

Undeclared Tool Scope

Medium
Category
MCP Least Privilege
Confidence
82% confidence
Finding

The skill explicitly describes creating questions.md, response.md, and meta.json, which implies file-write capability, but it does not declare any tool scope such as permissions or allowed-tools. In an agent setting, undocumented write capability weakens least-privilege controls and can lead to unintended file creation or writes outside the expected workspace if the runtime grants broader access than the user expects.

Content

No source excerpt is available for this finding.

Static analysis

No suspicious patterns detected.