Intent-Code Divergence
Medium
- Confidence
- 98% confidence
- Finding
- The guidance explicitly instructs the model to expose step-by-step chain-of-thought in the response, which requests disclosure of internal reasoning rather than a concise answer. That creates a policy and safety issue because users can extract internal deliberation patterns, and the instruction is embedded as model-facing guidance, increasing the chance it will be followed consistently.
