Description-Behavior Mismatch
Medium
- Confidence
- 97% confidence
- Finding
- The evaluator enables external LLM judging by default and is configured to send evaluation inputs to Anthropic. That means queries, expected answers, generated answers, and potentially source-derived content can leave the local environment without explicit opt-in, which creates a real data exfiltration and privacy/compliance risk in an evaluation tool.
