Hidden Instructions
High
- Category
- Prompt Injection
- Content
目前所有基于 HF Open-LLM-Leaderboard 发布的模型评分和排名都受到空白字符漏洞影响。lm-evaluation-harness(实际应用中最常用的 LLM 评估工具包)被发现存在存在该漏洞(如果前面存在空白字符,就会导致正确选项无法被正确选中)。 - Relative Adoption Metric (RAM)([atomproject.ai/relative-adoption-metric](https://atomproject.ai/relative-adoption-metric))
- Confidence
- 60% confidence
- Finding
- Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.
