Pith. sign in

Paper Citation Record · LEDGER

SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2403.01976.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.01976 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:10:20.605939Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

14
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 03694014-52e2-4d23-a59a-ca3afde7f50d · inbound

EarthSE: A Benchmark for Evaluating Earth Scientific Exploration Capability of LLMs cites this paper.

EarthSE: A Benchmark for Evaluating Earth Scientific Exploration Capability of LLMs SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:20.605939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:10:20.605939Z digest=sha256:cf882418d0e1dcff1f3cb036ff9589d9d19cefc1c29dbbf976c3d99174b829f2

Observation 95578303-0a4e-43f2-8890-0b52b5acb70c · inbound

Toward Scientific Reasoning in LLMs: Training from Expert Discussions via Reinforcement Learning cites this paper.

Toward Scientific Reasoning in LLMs: Training from Expert Discussions via Reinforcement Learning SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:32.807425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:32.807425Z digest=sha256:2e5b86a26c374e12d01bb885db5c7137281b203f6ad1ee81558f3a39f007199c

Observation f140799d-8d32-4e50-b37b-18fc74212e29 · inbound

Benchmarking Large Language Models on Homework Assessment in Circuit Analysis cites this paper.

Benchmarking Large Language Models on Homework Assessment in Circuit Analysis SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:50.503811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:50.503811Z digest=sha256:0e928e0171f01580a71de4af0c562fa75d0ba24ec3f918768253809e27983f1c

Observation d039c5ae-6c27-4e98-b3f5-843df3889575 · inbound

StatEval: A Comprehensive Benchmark for Large Language Models in Statistics cites this paper.

StatEval: A Comprehensive Benchmark for Large Language Models in Statistics SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T10:36:22.941355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:36:22.941355Z digest=sha256:0d28c1f83afbeea61da4cf30de5728ee31386ec1cae4e398b2f27f4e941cc22e

Observation 8254c6b1-b059-46ac-b680-e580821c9259 · inbound

Babel: Jailbreaking Safety Attention via Obfuscation Distribution Optimized Sampling cites this paper.

Babel: Jailbreaking Safety Attention via Obfuscation Distribution Optimized Sampling SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:08:11.073134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:08:07.648295Z digest=sha256:33866052840360efc7136f1078f6daca3329aa5ea1b279fd716740f8b21c5ea4

Observation a5d90bfe-a1c1-4f07-a20a-f4948062eedf · inbound

From Text to Discovery: How Large Language Models Are Reshaping Research Across Scientific and Humanistic Disciplines cites this paper.

From Text to Discovery: How Large Language Models Are Reshaping Research Across Scientific and Humanistic Disciplines SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T12:03:22.105331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:03:22.105331Z digest=sha256:55d50221e5d54862fc72f1f3a91eb1810066e9d960d4f8601fb803502a1c7845

Observation 7405d2ed-7030-4a64-bdc3-94aa3d15b7a9 · inbound

PeerCheck: Enhancing LLM-Generated Academic Reviews Towards Human-Level Quality cites this paper.

PeerCheck: Enhancing LLM-Generated Academic Reviews Towards Human-Level Quality SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T04:09:34.562797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T17:12:38.192534Z digest=sha256:a3f7a47edbe33d9eb9c1d0e2ca7aaf37c9596fe70912e158bce30d3943806c94