Vague Triggers
Medium
- Confidence
- 88% confidence
- Finding
- The manifest description says the skill should activate when the user mentions broad phrases like "compare model outputs," "create evaluation rubrics," or general concepts such as "automated quality assessment." These triggers are not tightly scoped and no exclusion conditions or negative examples are provided, which could cause the skill to activate in contexts beyond the intended advanced LLM-as-judge use case.
