Natural-Language Policy Violations
Medium
- Confidence
- 94% confidence
- Finding
- The skill hardcodes Chinese acknowledgement examples as the prescribed foreground reply pattern, which can override the user's language preference and cause confusing or inappropriate responses in multilingual environments. While not a direct security exploit, it can degrade trust, create miscommunication, and violate expected interaction boundaries when the user did not opt into Chinese output.
