Pith. sign in

Paper Citation Record · LEDGER

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification

As of 13 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 2 inbound Pith citation observations for arXiv:2412.16486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.16486 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:34:53.395062Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:34:53.395062Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T08:56:03.223978Z

Reference resolution

21 of 21 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2b4a5612-fefa-4835-83fe-7ca93fea0de0 · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.300608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.300608Z digest=sha256:88e3739c3a4e6b847af58eb685a3c7440f914eab64591ad4d0897af2e86ef0cf

Observation 0a61f1ed-4f5f-48d0-b814-84f148817bce · outbound

This paper cites Curran, Allyson Gallant, Helen Wong, Ca ther- ine Johnson, Alannah Delahunty-Pike, Lynora Saxinger, Derek Chu, Jeannette Comeau, Trudy Flynn, Julie Clegg, and Christopher Dye.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Curran, Allyson Gallant, Helen Wong, Ca ther- ine Johnson, Alannah Delahunty-Pike, Lynora Saxinger, Derek Chu, Jeannette Comeau, Trudy Flynn, Julie Clegg, and Christopher Dye

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:34:54.130939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T10:34:53.305841Z digest=sha256:af21f320f317c326ebd0367232fb8b3468813bf3c3d6e7b1d2329cef20109a57

Observation 2b91d2f8-cf18-46df-868e-5d3ad7594728 · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.315933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.315933Z digest=sha256:41ee52c4809cef2fefb7ca97bc87577341ecb775ddd26d1bd0a9ccc99be23e69

Observation 3f85c2db-2ca2-4863-a4f8-0108361af0cb · outbound

This paper cites Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.320762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.320762Z digest=sha256:d295d2a06c56aa8a8e95fe6d02211f0cfb3f69267a8208285cc9826a5608c61f

Observation d217ea4c-1bd7-46cc-9e9b-6403611b8e06 · outbound

This paper cites FakeGPT: Fake News Generation, Explanation and Detection of Large Language Models.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification FakeGPT: Fake News Generation, Explanation and Detection of Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.325668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.325668Z digest=sha256:1b4340954ca0698216c31335fb01f500aeacbc6d143058df13fcf554bec5a63b

Observation faea6f9f-8a41-4a50-b5e2-4a1d6b8710e9 · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.330531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.330531Z digest=sha256:694032c2312db7d004cc29f7956597b21079e91929205fb1c56d2fed1c04fb4f

Observation ae39ff4d-f4a7-4ba6-83c0-84d99d0e5fb8 · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.335365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.335365Z digest=sha256:83bdf4625012f83e61e991de32bd41503eff2ca15538c639d432247c9594ffcf

Observation e836cf94-1e7b-4444-bbed-f169f5ad633d · outbound

This paper cites Large Language Models Understand and Can be Enhanced by Emotional Stimuli.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Large Language Models Understand and Can be Enhanced by Emotional Stimuli

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.344478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.344478Z digest=sha256:98f574a0bdaba2eaa928bfda9128a1288addb4dfd000283c4b5cf623a79e64bf

Observation 7352865a-202b-4cd5-859a-9433e09add22 · outbound

This paper cites Mancenido, and Huan Liu.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Mancenido, and Huan Liu

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.348895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.348895Z digest=sha256:97ab82faa7fb3426101b8eb069b530ba5867ce5cc72f916190672e313e3c255a

Observation f382df45-2f07-45cd-bc17-54c00597d748 · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.353151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.353151Z digest=sha256:58e83ec5c6999dbfee09145283d627964afbe80fdb4d85dc487a6c38ce4b9c9d

Observation 137f70d7-0622-46c2-96d0-fc753b2f10ac · outbound

This paper cites COVID-Fact: Fact Extraction and Verification of Real-World Claims on COVID-19 Pandemic.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification COVID-Fact: Fact Extraction and Verification of Real-World Claims on COVID-19 Pandemic

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.357296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.357296Z digest=sha256:5046eaecf3595c13b33ce32a87d151c2ea0c8d3e638254a0e144cb36a706fdf3

Observation 498bbca0-0660-41db-8486-0ff00db504d7 · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.363264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.363264Z digest=sha256:3d6952f4f8c56766a4ddcf792555eac5f995f69eda78f312113b5d271815932b

Observation ee4243a0-52c6-4e30-819d-106eb1a3e2cc · outbound

This paper cites Text Classification via Large Language Models.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Text Classification via Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.367830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.367830Z digest=sha256:58122ce609f8fb7f9f580afe9e06ce0e9df7aeee5aa2dec1580cdce29db55781

Observation c05669c0-e568-413b-8a74-e822d0467f74 · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:34:54.091871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T10:34:53.372869Z digest=sha256:c00dfd1fcf2ca0a30ffb078d178bbe09d912469e277e2444ae0e11c1ed67faa7

Observation 0b0defab-611f-4f16-9a12-76b490e86ecc · outbound

This paper cites Claim Detection in Biomedical Twitter Posts.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Claim Detection in Biomedical Twitter Posts

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-11T10:34:53.634818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T10:34:53.381424Z digest=sha256:92f381f11f66d196325148dde51f6ffbd7a8f98213f33361bd14f11912d0bae0

Observation bab02088-f1a0-452f-a454-c9e0ce360a26 · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.386105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.386105Z digest=sha256:5d6ed7d3fef954231e9c177d664c38081f5e0c446c0892e9dbc39c6f59704844

Observation d4a924af-f5d8-4d3c-acde-aca9c720938b · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.390723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.390723Z digest=sha256:14606043be85b50e2d4f9b837eec0545b9bb1746aa5d1d47cf895fa8a5edde12

Observation e031025b-3683-4c4a-a802-cbbe0f575fe4 · outbound

This paper cites Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.395062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.395062Z digest=sha256:4a06f889ba44aa5c702f5a1c8f5b8a80db94c83d14e235c5939de0f413c2a124

Observation 7c7a64f6-3897-47d7-8f41-54bbbec8dce0 · outbound

This paper cites In Proceedings of the 34th International Conference on Neural In- formation Processing Systems (Vancouver, BC, Canada) (NIPS’20).

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification In Proceedings of the 34th International Conference on Neural In- formation Processing Systems (Vancouver, BC, Canada) (NIPS’20)

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:34:54.106725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T10:34:53.339900Z digest=sha256:8b1ad12ddbe80a64c572f8a0923e6c17164fff963374e780b77de07973116bdf

Observation 9f79e981-ce1c-43c1-8544-778b67348c8e · outbound

This paper cites Philosophical Transactions of the Royal Society A: Mathematical, Physical and Engineerin g Sciences 381, 2257 (2023), 20230133.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Philosophical Transactions of the Royal Society A: Mathematical, Physical and Engineerin g Sciences 381, 2257 (2023), 20230133

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.310694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.310694Z digest=sha256:984af06a3e4203d256a7827c5ef448fcb7b5f684ade50c2bc9887d9069ada4ac

Observation bc71bdaa-c869-4b44-984d-873edc47978d · outbound

This paper cites an unresolved cited work.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Unresolved cited work

Reference 4734

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.377167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.377167Z digest=sha256:25173186e0e51c82b9a8c81311960dff185496703bb3db9767da939203999893

Pith citing papers

Observation e031025b-3683-4c4a-a802-cbbe0f575fe4 · inbound

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification cites this paper.

Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:53.395062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:34:53.395062Z digest=sha256:4a06f889ba44aa5c702f5a1c8f5b8a80db94c83d14e235c5939de0f413c2a124

Observation e18b32e8-a338-460a-93ba-9caeeb59764e · inbound

Uncertainty-Aware Web-Conditioned Scientific Fact-Checking cites this paper.

Uncertainty-Aware Web-Conditioned Scientific Fact-Checking Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:56:03.226118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T16:23:17.293800Z digest=sha256:e502be6c4a11929105334d9eeb7fe833b1d675fc62e5fae649d61796b1012060