Natural-Language Policy Violations
Medium
- Confidence
- 91% confidence
- Finding
- The skill hard-codes a Chinese-translation workflow as the default interpretation of the task, which can override or bias the agent away from the user's actual language and output preferences. This is not code execution or data exfiltration, but it can cause unauthorized transformation of content, fidelity loss in summaries/translations, and user-intent misalignment in downstream deliverables.
