Pith. sign in

Paper Citation Record · LEDGER

PromptBench: A Unified Library for Evaluation of Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2312.07910.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.07910 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:35:42.930625Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

13
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation acac4324-7306-4d43-9b7a-3ea2b6ec88c4 · inbound

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models cites this paper.

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 62

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:08:05.558387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T06:08:05.386345Z digest=sha256:ea9b56c7522ded84f38f040ed96e44efbf3efb63cc67c6b192efa4af361cb366

Observation d98d3097-f199-4160-b3fa-6426e57924ce · inbound

The Science of Evaluating Foundation Models cites this paper.

The Science of Evaluating Foundation Models PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-07T23:35:42.930625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:35:42.930625Z digest=sha256:4728cf2036fa688f052a6f6f41a7150a8d790b688d54ff464eb4c53885a6daf8

Observation e3b17ae8-85ee-4fe3-a048-05050b577aba · inbound

FLASH-D: FlashAttention with Hidden Softmax Division cites this paper.

FLASH-D: FlashAttention with Hidden Softmax Division PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:44.481313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:44.481313Z digest=sha256:56bbfc4ad4c638926c0648a19f8557cbf66dbe78d620cb285f867afa0b02c7a9

Observation e66f864b-18c8-497f-86bf-0902a7c6d8fd · inbound

Low-Cost FlashAttention with Fused Exponential and Multiplication Hardware Operators cites this paper.

Low-Cost FlashAttention with Fused Exponential and Multiplication Hardware Operators PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:48.583082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:48.583082Z digest=sha256:28256a388e9b1754d4226504a580b49590bad358c7dfa83d26522d21be573532

Observation 4c02b2ff-2e11-4657-82bd-5b6b40fc5fa1 · inbound

Safe-Child-LLM: A Developmental Benchmark for Evaluating LLM Safety in Child-LLM Interactions cites this paper.

Safe-Child-LLM: A Developmental Benchmark for Evaluating LLM Safety in Child-LLM Interactions PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:32:16.304711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T09:31:14.739478Z digest=sha256:de8a074b8bd4c98a76b4eb02eb2cf014b89c6d3c79a4c00a2a6c2676ab45faf5

Observation 6472e16d-5097-4478-ba43-bbfb3f9c48e8 · inbound

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing cites this paper.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:53.310912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:53.310912Z digest=sha256:1d7fcecbc799c419163ab76e8a082649fa85949490e8b499b5ddcc5f12022975

Observation 0c1af361-7a19-48a6-bc2f-77de003b6a73 · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.237976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.237976Z digest=sha256:97d26cd8cb2555bbb8baab35dedf68ba7e5a6124268d65c302f738ef93a7a93b

Observation 05aaff4a-e3e7-4ee0-ba8a-84952cae68f1 · inbound

Characterizing Fitness Landscape Structures in Prompt Engineering cites this paper.

Characterizing Fitness Landscape Structures in Prompt Engineering PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:27:28.018193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:27:28.018193Z digest=sha256:15ec13d3c2b1115a74c0668179e2f36c510f893f3f68872495c43dafaeedb699

Observation 7366d3af-7c6c-43b4-8f03-19f983a2ae43 · inbound

Measuring Behavior Portability in Large Language Models cites this paper.

Measuring Behavior Portability in Large Language Models PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-26T09:09:16.242891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T08:59:20.724868Z digest=sha256:0c4f6e218e7f9fbf24e2873a61b311d6388a0456f994ce9293278ab5cff8b77a

Observation ab524a00-2979-4d1d-8413-3f01fdfb7fa5 · inbound

Simplicity Paradox: Debunking myths about prompting and datasets for LLM evaluation cites this paper.

Simplicity Paradox: Debunking myths about prompting and datasets for LLM evaluation PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T14:42:46.785295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:42:46.785295Z digest=sha256:b0a1c27c3b05d1965f90c8c068580cef3b7591d32905a24964c7c2c1368d5880

Observation 082da28f-b19c-4774-b6d8-a87b3d697200 · inbound

MolecularCanvas: LLM-assisted Small-Molecule Drug Discovery via Structure-Guided Constraints cites this paper.

MolecularCanvas: LLM-assisted Small-Molecule Drug Discovery via Structure-Guided Constraints PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T04:16:59.551189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:16:59.551189Z digest=sha256:bb7d45dd51e5519ff882d1bb94f2f4f07f5d0d090ceff0fca5c75304315d9de3