Pith. sign in

Paper Citation Record · LEDGER

Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2306.11270.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.11270 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T22:54:48.347005Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:58:57.677920Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7efef2b2-85b0-4390-957a-33ec2e23c547 · inbound

Beyond Prompt Content: Enhancing LLM Performance via Content-Format Integrated Prompt Optimization cites this paper.

Beyond Prompt Content: Enhancing LLM Performance via Content-Format Integrated Prompt Optimization Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T22:54:48.347005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:54:48.347005Z digest=sha256:411c76e3e19c70edc6ca842e4c2d14b1768afa6c85d427c51e879908b80557aa

Observation 99af6741-ddb7-48db-8fce-5ec9a5da0b7e · inbound

Investigating the Robustness of Retrieval-Augmented Generation at the Query Level cites this paper.

Investigating the Robustness of Retrieval-Augmented Generation at the Query Level Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:55:54.909052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:55:54.909052Z digest=sha256:348db0efff3e237eaef7295f63e2a5237a5de84a1f7b56ff469a18004c617ec5

Observation 65904cfc-6982-4dd2-b3c3-78d6f22486b2 · inbound

PlotTwist: A Creative Plot Generation Framework with Small Language Models cites this paper.

PlotTwist: A Creative Plot Generation Framework with Small Language Models Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T18:06:40.507466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:06:40.507466Z digest=sha256:7294ae8b1cc3704348eace34b4d75139ec7a066275c9cf137e6fd6347959c8ef

Observation 7443e39b-e8e7-4138-9734-159fa2c6c416 · inbound

Compared to What? Baselines and Metrics for Counterfactual Prompting cites this paper.

Compared to What? Baselines and Metrics for Counterfactual Prompting Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 164

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:56:27.679391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-09T19:02:46.991897Z digest=sha256:c947d37cf2dd5da6508dcb41aac0958dcf13a8fcef2d89f66a9f668aadd599af

Observation bb2d9f26-1830-4dd4-af72-4cb87a97d3f6 · inbound

Towards Context-Invariant Safety Alignment for Large Language Models cites this paper.

Towards Context-Invariant Safety Alignment for Large Language Models Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.890191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T05:36:12.562807Z digest=sha256:3d34c9b7475411e6b51520b3d9124893755db20c9147432ba54dc7dfbbae4290

Observation e58e76aa-9eb0-4a23-ad73-5cbda748458f · inbound

Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs cites this paper.

Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:23:13.274114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:14:22.444020Z digest=sha256:2422998ed9d63590a1e80d5394ea5327db38182465c9ee63dc1de5e58f827f89

Observation c103f8f5-6aaf-4a19-8e22-f6956c057a90 · inbound

Prompt Perturbation for Reliable LLM Evaluation over Comparison Graphs cites this paper.

Prompt Perturbation for Reliable LLM Evaluation over Comparison Graphs Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:58:57.679681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T01:03:20.261268Z digest=sha256:4fac6f94b5c26603d1c4777a6abbf236b10548fab9078d8bad3f51e9399531a2

Observation 6d28b749-88ac-4093-a4af-38eab487840a · inbound

The Powerless Noise: How Experimental Settings Shape the Reported Power of Noise cites this paper.

The Powerless Noise: How Experimental Settings Shape the Reported Power of Noise Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T01:10:52.578408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:10:52.578408Z digest=sha256:854c70a48ccbae0a8dd5921cc850840b0729b254a0b0964619fc002d983fca76

Observation 688187b5-de1e-443f-b59d-069303e3f4ca · inbound

The Powerless Noise: How Experimental Settings Shape the Reported Power of Noise cites this paper.

The Powerless Noise: How Experimental Settings Shape the Reported Power of Noise Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-13T07:04:18.645464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:04:18.645464Z digest=sha256:cc73c1af0056cbf03d554afa08622addc7930c3d4ae8d64d2a371ffc775df33f