Pith. sign in

Paper Citation Record · LEDGER

PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.19740.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.19740 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T13:40:46.709807Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:06:35.205226Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a9011def-dae2-4f1d-87ff-a979333e27b6 · inbound

LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient cites this paper.

LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T18:09:01.639375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:09:01.639375Z digest=sha256:d5e2ef3f032c438d1b59bb61edf11a1a6e86e026f8176a5d9c06c24610170ea0

Observation 9a3bb878-238b-4286-a246-8037bd818d26 · inbound

Governing AI Beyond the Pretraining Frontier cites this paper.

Governing AI Beyond the Pretraining Frontier PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:46.709807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:46.709807Z digest=sha256:494d8fd9fcf1e278cdeceb956ef25f38c211295ab25323c950cbaaac416dcd81

Observation 53df4759-7c2d-4fb4-9e65-93c92e196499 · inbound

AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data cites this paper.

AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:01.719577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:01.719577Z digest=sha256:72e63f60770ae5e55ca4d6671be401d5e78bbd6e835a5dd80989967a577726e6

Observation 53967675-f847-4a91-9882-90e335104bdf · inbound

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules cites this paper.

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:31:25.892228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-12T01:00:13.017290Z digest=sha256:41529503256db8cf33fe118eab8be1707f2044768dfbe63f7de9eee043ed9cd7

Observation c399ec89-b0ad-42a5-a561-b12ffff554c6 · inbound

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks cites this paper.

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:06:35.206717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-28T09:27:30.923556Z digest=sha256:b5f9328dccb2c986974919b3517affcf04d74f5cecff97fd49bf4ac1b175f793