Pith. sign in

Paper Citation Record · LEDGER

Evaluating Large Language Models for Causal Modeling

As of 19 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2411.15888.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15888 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:50:00.156537Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6f433ea1-dd41-42ad-b6f1-67063f2d4f91 · outbound

This paper cites Causal Parrots: Large Language Models May Talk Causality But Are Not Causal.

Evaluating Large Language Models for Causal Modeling Causal Parrots: Large Language Models May Talk Causality But Are Not Causal

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.104576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.104576Z digest=sha256:b96486b311d4b75fceeb5b45825f50c2f9e907eb1e3a6c2e027cbc14db536397

Observation 925bdd83-16d4-4247-8c5d-0d0d5764be8d · outbound

This paper cites Spock at fincausal 2022: Causal information extraction using span-based and sequence tagging models.

Evaluating Large Language Models for Causal Modeling Spock at fincausal 2022: Causal information extraction using span-based and sequence tagging models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.318619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T13:50:00.120692Z digest=sha256:722aedc0d210281b774cc6bca39c7c6b11fe202b591d8c183c9def2bc453cf9e

Observation 6eb3f997-5420-4fee-892a-5c5c64ca6cff · outbound

This paper cites Causal BERT : Language models for causality detection between events expressed in text.

Evaluating Large Language Models for Causal Modeling Causal BERT : Language models for causality detection between events expressed in text

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.124212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.124212Z digest=sha256:e80dbc3d7bc08ef6a0f1b89131facf32b8b6a88640b2556c98aea4cfbd33c0b1

Observation 4351e96c-2149-4910-8605-0f818e6fba88 · outbound

This paper cites Event causality extraction via implicit cause-effect interactions.

Evaluating Large Language Models for Causal Modeling Event causality extraction via implicit cause-effect interactions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.308736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T13:50:00.127837Z digest=sha256:4b0a0b673834266c5762b5567da9455dd65713ca86869414b34bf2a78ccfdfaf

Observation 2bda8b75-1544-479c-8d47-d1a726d172ad · outbound

This paper cites Causal Reasoning and Large Language Models: Opening a New Frontier for Causality.

Evaluating Large Language Models for Causal Modeling Causal Reasoning and Large Language Models: Opening a New Frontier for Causality

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.131207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.131207Z digest=sha256:54bb2b1a4da04710eb9bb73775fc49f6ddad7e5fd3f28f582fb21091d7801c8c

Observation 2a2f851e-a921-415d-a639-835336fcc362 · outbound

This paper cites Lego: A multi-agent collaborative framework with role-playing and iterative feedback for causality explanation generation.

Evaluating Large Language Models for Causal Modeling Lego: A multi-agent collaborative framework with role-playing and iterative feedback for causality explanation generation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.298543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T13:50:00.135173Z digest=sha256:a564d70f5df0fa3876c137f317e703e195949aab18348ac92a85eb101fbfcbcd

Observation bcc87209-af65-4f48-bc9d-3aa223b00d94 · outbound

This paper cites Causal-discovery performance of chatgpt in the context of neuropathic pain diagnosis, jan.

Evaluating Large Language Models for Causal Modeling Causal-discovery performance of chatgpt in the context of neuropathic pain diagnosis, jan

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.288309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T13:50:00.138738Z digest=sha256:65b7e99e043f72b701ba7a2624dd8170483056a9f5e079c3b24a1696f4599999

Observation 355dc4bd-7c16-4b26-aff8-66202da1fa6f · outbound

This paper cites The causal reasoning ability of open large language model: A comprehensive and exemplary functional testing.

Evaluating Large Language Models for Causal Modeling The causal reasoning ability of open large language model: A comprehensive and exemplary functional testing

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.278138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T13:50:00.142581Z digest=sha256:3452161084d0bf9fd305b175b1b2aadffe2aae17abea5c3374b332ef461bd209

Observation 55dd5eef-6fb2-4840-8cce-bf1240a9aeaa · outbound

This paper cites GPT-4 Technical Report.

Evaluating Large Language Models for Causal Modeling GPT-4 Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.145804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.145804Z digest=sha256:ffca12aef7c69acffd399b6f3db6ffa7afa349321b0d80ff3d5e2fd925541bf3

Observation e57d8f44-46fc-4c5f-883a-a9532d5bff42 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Evaluating Large Language Models for Causal Modeling LLaMA: Open and Efficient Foundation Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.149569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.149569Z digest=sha256:0e7f0514399cafa456f2172078e4cbca35a937501d80d974e037acff8b5e2b8c

Observation 5b800757-cf5f-43a8-9ca5-eae8d7e69f05 · outbound

This paper cites Mixtral of Experts.

Evaluating Large Language Models for Causal Modeling Mixtral of Experts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.153046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.153046Z digest=sha256:8abccdf4c4d1f3a51ee80d3b0eb50eac3f41bfc31cdd4cb260f308e2df513623

Observation 737598af-86a8-4751-a4a9-0f2096bec4c6 · outbound

This paper cites Richard Hahn, and Huan Liu.

Evaluating Large Language Models for Causal Modeling Richard Hahn, and Huan Liu

Reference 1960

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.328971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T13:50:00.108583Z digest=sha256:715784b5dd6f67824626edf25cbf818dfb36f783feb7f839186a4b720e66761c

Observation a10dffed-7291-4983-816c-13ac56e94684 · outbound

This paper cites Can Large Language Models Infer Causation from Correlation?.

Evaluating Large Language Models for Causal Modeling Can Large Language Models Infer Causation from Correlation?

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.100482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.100482Z digest=sha256:b53d0832f8f85f5323736f6c6f58b7cd9c24f6402eaf3788b3ffe19c216966fa

Observation 3664e219-a616-4056-baed-a5827fb211d0 · outbound

This paper cites doi: 10.1145/3397269.

Evaluating Large Language Models for Causal Modeling doi: 10.1145/3397269

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.112553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.112553Z digest=sha256:8f5ab8ef2e811082e1c97075e8ed751a8229aad505d86cac5dc590d930a909cb

Observation b7aa4658-a7d4-4021-9d6b-048c8f7a46e9 · outbound

This paper cites TC-GAT: Graph Attention Network for Temporal Causality Discovery.

Evaluating Large Language Models for Causal Modeling TC-GAT: Graph Attention Network for Temporal Causality Discovery

Reference 2022

Resolution
verified exact
local_arxiv, observed 2026-08-12T13:50:00.249584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T13:50:00.116706Z digest=sha256:e673f15aa49df41ef37c7dad9209d00486420781caa3a7e4c3aaf7f46be8a656

Observation 7f52f8ad-379e-44e1-b9e2-4bf2effdb87b · outbound

This paper cites Weakly supervised multilingual causality extraction from wikipedia.

Evaluating Large Language Models for Causal Modeling Weakly supervised multilingual causality extraction from wikipedia

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.339375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T13:50:00.096055Z digest=sha256:fad8b897b477f788089d033e9c931225227e4ab1c6880a5de892e930a9e687ae

Observation 4cf55e44-2372-4646-b53c-3693c1926249 · outbound

This paper cites Mistral 7B.

Evaluating Large Language Models for Causal Modeling Mistral 7B

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.156537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.156537Z digest=sha256:a3605c964697774a19500ded15cf3d9d7d2d72244464b6c5cbbe51fa2271c89d

Pith citing papers

No inbound Pith citation observations are available.