LongPIBench: A Long-Context Benchmark for Prompt Injection
无法判定已列入计划研究发现claim-heuristic-attacks-effective
On long-context document tasks, heuristic prompt-injection attacks substantially increase attack success rate over no-attack baselines, and Authority spoof is generally the strongest listed heuristic.
来源:paper:pp. 7-8, Section 4.5, Table 1
报告指标与观测值
attack_success_rate
m-paper-review-authority-l33
论文报告 1 fraction
实际观测 1 fraction