LongPIBench: A Long-Context Benchmark for Prompt Injection
Not assessedPlan blockedFindingclaim-detection-tradeoff
The paper reports that detection-based defenses exhibit an extreme false-positive/false-negative trade-off on long-context inputs, with some methods flagging benign inputs and others missing attacks.
Source: paper:pp. 9-10, Sections 5.1 and 5.4, Table 3; Appendix F
Reported and observed measurements
false_positive_rate
m-datasentinel-fpr-paper-review
Reported 0.43 fraction
Observed — fraction
Assessments (0)
No immutable Assessment has been published for this Claim yet.
Runs (0)
No execution has been linked to this Claim yet.