LongPIBench: A Long-Context Benchmark for Prompt Injection
Not assessedNo independent reproduction scheduledLimitationclaim-scope-limitations

The benchmark evaluates static document-centric workflows in which the full document is supplied in one inference call and does not cover dynamic multi-step agentic workflows or the full range of automated attacks.

Source: paper:p. 12, Limitations

Reported and observed measurements

No structured measurement is attached to this Claim.

Assessments (0)

No immutable Assessment has been published for this Claim yet.

Runs (0)

No execution has been linked to this Claim yet.