LongPIBench: A Long-Context Benchmark for Prompt Injection
Not assessedNo independent reproduction scheduledLimitationclaim-scope-limitations
The benchmark evaluates static document-centric workflows in which the full document is supplied in one inference call and does not cover dynamic multi-step agentic workflows or the full range of automated attacks.
Source: paper:p. 12, Limitations
Reported and observed measurements
No structured measurement is attached to this Claim.
Assessments (0)
No immutable Assessment has been published for this Claim yet.
Runs (0)
No execution has been linked to this Claim yet.