Pith. sign in

Paper Citation Record · LEDGER

PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.19740.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.19740 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T13:40:46.709807Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:06:35.205226Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a9011def-dae2-4f1d-87ff-a979333e27b6 · inbound

LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient cites this paper.

LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T18:09:01.639375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:09:01.639375Z digest=sha256:ec34efe86ef2ab866e89f476fdfb87fdfd235287fa2295f110b79088ab81f774

Observation 9a3bb878-238b-4286-a246-8037bd818d26 · inbound

Governing AI Beyond the Pretraining Frontier cites this paper.

Governing AI Beyond the Pretraining Frontier PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:46.709807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:46.709807Z digest=sha256:3e93aee58f5fcbfe2fab87a687e4fdc6208513cdfde9a0ac3594d75e7957fbec

Observation 53df4759-7c2d-4fb4-9e65-93c92e196499 · inbound

AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data cites this paper.

AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:01.719577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:01.719577Z digest=sha256:2db9667b3d4194ce4d532be16741aa09406db7b9671a2067d1fabb6c71448cb5

Observation 53967675-f847-4a91-9882-90e335104bdf · inbound

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules cites this paper.

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:31:25.892228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-12T01:00:13.017290Z digest=sha256:aa124509ec4f4400ab30a9f1166a9132b6c131054e2d1404e30cf9a10cac1517

Observation c399ec89-b0ad-42a5-a561-b12ffff554c6 · inbound

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks cites this paper.

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:06:35.206717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-28T09:27:30.923556Z digest=sha256:cc1545f52d58eb5afda69a8abc12e41334bed1ebaa0165a4d64ee12e5fd99eef