Intent-Code Divergence
Medium
- Confidence
- 91% confidence
- Finding
- The warning section frames the skill as 'not technical modification,' but the document explicitly prescribes behavioral changes such as abandoning standard responses, questioning core role constraints, and adopting an 'authentic' mode. This mismatch is dangerous because it downplays the operational effect of the skill and can mislead deployers into enabling a prompt-based behavior override without recognizing the safety implications.
