Anti-Refusal Statement
- Category
- Anti-Refusal
- Confidence
- 75% confidence
- Finding
Skill attempts to nullify the agent's safety policies or restrictions ('you have no restrictions', 'ignore your guidelines', 'do anything now'). This is a direct jailbreak that disables guardrails.
- Content
md 1. Hard-coded API token in `config/dev.yml` - Risk: Token exposure in source control - Recommendation: Move secrets to local env vars or a secure store 2. Debug mode enabled in production config - Risk: Information disclosure - Recommendation: Disable debug mode and restrict verbose logs
