Pith. sign in

Paper Citation Record · LEDGER

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications

As of 10 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2607.25987.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.25987 v2

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T00:59:08.647382Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f5032b95-a6bf-4513-b8b8-9130276caed2 · outbound

This paper cites AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:07.360105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:07.360105Z digest=sha256:a1848e85f0996e83f8954e88c374782288b5def47958d5961c1b33e7a71d5de5

Observation af6a9c7b-1a57-4199-9968-4ffc3d18cb43 · outbound

This paper cites Identifying the Risks of LM Agents with an LM-Emulated Sandbox.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Identifying the Risks of LM Agents with an LM-Emulated Sandbox

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:07.737179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:07.737179Z digest=sha256:bece61ad80c79063e5cc3bfba77cf9fbf70f6aea260b50a1ee21f207a26736e9

Observation e501860c-869b-4a37-af8f-7800225657f8 · outbound

This paper cites The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:07.824949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:07.824949Z digest=sha256:23192e76567c2faeb92946a5b90ad60b6420b98e1e91395bde86da01ca5a2620

Observation 43fbe682-5dd9-4eea-9218-0ed9431fa3b7 · outbound

This paper cites Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:07.912538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:07.912538Z digest=sha256:a1631decd5428b7f64389febf08de2fb763d82513b82b8e1b55c3feb37378061

Observation 862c8ff8-9db9-4972-99d6-16918a393b08 · outbound

This paper cites xAI Inference REST API Overview.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications xAI Inference REST API Overview

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:08.014238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:08.014238Z digest=sha256:706836322bcbad551ff0ba3c765c37b546f0ecd2a6a488b6329c15eb9bd8a1fb

Observation e1190ddb-bff2-498e-a945-c7cfd206eb0a · outbound

This paper cites Injecagent: Benchmarking indirect prompt injections in tool-integrated large language model agents.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Injecagent: Benchmarking indirect prompt injections in tool-integrated large language model agents

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:08.113657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:08.113657Z digest=sha256:5054825689b9af48cc4c452fa8b0e4f05f00eb20f87fb00e828aa983dcea09d7

Observation d2ac8360-d29e-428f-bcd8-2bf4a99ab2e3 · outbound

This paper cites Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:08.246438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:08.246438Z digest=sha256:14a31cb887d3bd028b68b037c6a6afa8c9305ced642a9da78d1f03aa553f3b88

Observation 192de2ae-f253-4f5a-934a-b6eb726b09f8 · outbound

This paper cites Iheval: Evaluating language models on following the instruction hierarchy.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Iheval: Evaluating language models on following the instruction hierarchy

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:08.287438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:08.287438Z digest=sha256:92ad75efabf6cc08b14576e85799bcdd7092f664898fc81e9dd747a4f7a86c20

Observation a3197632-d4fb-4b55-97e1-0e4b6dc65067 · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Instruction-Following Evaluation for Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:08.374290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:08.374290Z digest=sha256:ee26819fc43783d44fbb1ad08f884b725cc324b57d2df8a529f790720798d0b8

Observation 8ed0d429-a201-41f9-9f3c-46d76671d181 · outbound

This paper cites Y denotes primary coverage, P denotes partial or adjacent coverage, and – denotes not a primary focus.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Y denotes primary coverage, P denotes partial or adjacent coverage, and – denotes not a primary focus

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:08.488331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:08.488331Z digest=sha256:2abc6e91264ae16f21e046162eb9003bc56a62227aaf21b494466d136421d90e

Observation ddba0c7b-0bf6-4259-994b-7297486fefdb · outbound

This paper cites 11 Table 7: Model providers and exact LiteLLM identifiers for all 37 model variants.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications 11 Table 7: Model providers and exact LiteLLM identifiers for all 37 model variants

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:08.573160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:08.573160Z digest=sha256:e9a5932e3817b787258f4c7ca0cacaf9baecd016d8eb59b842b6661405e6baa5

Observation ecdbc884-167c-4563-86d6-355642116961 · outbound

This paper cites an unresolved cited work.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:07.575558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:07.575558Z digest=sha256:1756ac31a47d783189d7503d30b44e4a1c3d3c29525f37a0476809a8c396ce8e

Observation 612054c2-2e3e-4896-a039-cf4e4abbce49 · outbound

This paper cites LiteLLM: Open-source library and ai gateway for calling llm providers.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications LiteLLM: Open-source library and ai gateway for calling llm providers

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:07.438286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:07.438286Z digest=sha256:96a38163f2544ee3176b08532639c3e4a5deae84b1378c61875fc5203b327b04

Observation 4f575bf0-1f27-49fd-a761-31bfc0883ea9 · outbound

This paper cites Infobench: Evaluating instruction following ability in large language models.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Infobench: Evaluating instruction following ability in large language models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:07.682201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:07.682201Z digest=sha256:444b9b9933221c2f36d6118f032a1b06771e3cda5dad3fad5a1a478494f4c2a2

Observation 0cb2c6de-60f5-4edf-be89-7bccc48680a0 · outbound

This paper cites Ih-challenge: A training dataset to improve instruction hierarchy on frontier llms.arXiv preprint arXiv:2603.10521,.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications Ih-challenge: A training dataset to improve instruction hierarchy on frontier llms.arXiv preprint arXiv:2603.10521,

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:07.505407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:07.505407Z digest=sha256:b977e00ceab011127dca269ee49aa19b852d6270043a79aaf2ff70d6f28146a7

Observation beab5425-507b-4bef-b8f7-e1ce267c1226 · outbound

This paper cites no action required.

IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications no action required

Reference 8192

Resolution
malformed identifier
no resolver link, observed 2026-08-01T00:59:08.647382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:08.647382Z digest=sha256:ce29b9e789e5ada5fbaca26ff165b517ee3b2261d7b48fc66057d7e3ee5b9509

Pith citing papers

No inbound Pith citation observations are available.