Pith. sign in

Paper Citation Record · LEDGER

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs

As of 10 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2502.02896.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.02896 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:44:41.145054Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact7
  • verified fuzzy12
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c02d07c5-c1bc-4984-8422-a20e7cc26d16 · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:40.967981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:40.967981Z digest=sha256:740bd4e52e718af19f527762e8732a59c777948d5b548ae373145577a9c1709b

Observation e7081593-2b30-49f0-9104-a2e77789caf2 · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:40.973429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:40.973429Z digest=sha256:6d4890b6b3761fdc1f7170dd5e024b37484baaf2bfdbcef3fce02ca641dafaf2

Observation 1f20b524-c4c3-435e-965e-b1520a7ec2a4 · outbound

This paper cites Koutsiana, J.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Koutsiana, J

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:40.978142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:40.978142Z digest=sha256:3124d206024848cfa84a134fe3641f8fecaeda56c7981546982b86f8e9ccc9d6

Observation 4b311f49-816b-4cd7-9484-79b28e3931e6 · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:44:41.933813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:40.982854Z digest=sha256:804cdfb07340a1f925edee7e61bda77bd9828b53148c6047529286da30a3c6f3

Observation c6bc7544-0ae3-4e2e-b219-9c4cfd4d64f6 · outbound

This paper cites A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:40.987851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:40.987851Z digest=sha256:365f044a16f3b87817ec45fcccc88ad834a4e54bec32cb108427c057cbc6e81b

Observation 4a1a8053-efa9-4829-9497-0f6a199a8725 · outbound

This paper cites Mickus, E.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Mickus, E

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.918875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:40.993396Z digest=sha256:6701741532c0340999d62f995bf270b0c37bbeb552a9218b75b76c97e481ff6b

Observation 65ef6303-f4f0-4fe3-9d56-80217d083c21 · outbound

This paper cites WildHallucinations: Evaluating Long-form Factuality in LLMs with Real-World Entity Queries.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs WildHallucinations: Evaluating Long-form Factuality in LLMs with Real-World Entity Queries

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:40.999318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:40.999318Z digest=sha256:691de20452988919f9b1afa5fa2c889e730d2b06a7952514ed06bf8bd41524d5

Observation 210837a3-5568-4ad4-99f8-a8b503349d7d · outbound

This paper cites Language Writ Large: LLMs, ChatGPT, Grounding, Meaning and Understanding.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Language Writ Large: LLMs, ChatGPT, Grounding, Meaning and Understanding

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:44:41.473023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.005227Z digest=sha256:b75a1717b2d6cd691c625dc7f9738d7ae4fb878938cdf3884ae343106b03b122

Observation 74ac46a2-0ad4-433a-8a05-19e091f9047c · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:41.010402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:41.010402Z digest=sha256:5cdf178044285dc1c5b031af56135e4ac60af4c41d82ef2b1c30a92caab0de9b

Observation 262e9596-d95c-439e-95e5-7bca45c13c30 · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:44:41.903803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.015461Z digest=sha256:42f23e90b37b09a84212cd6b79a3500d1013550caa7886f21a2e17deaa478e5b

Observation 0bc3fd8d-c396-4d4e-b1b1-7877b3797bff · outbound

This paper cites Language Models as Knowledge Bases?.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Language Models as Knowledge Bases?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:41.020316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:41.020316Z digest=sha256:8132bb3978ea4773db7602eed1ea5a974b6eb0fb233bf85fc75009a369ecd703

Observation 5e48f049-bb9f-4863-977e-e89609d0b630 · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:44:41.888553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.025454Z digest=sha256:0409b737ad14b2d0dd34cf612ee179398f9f02408475ebb3e06346b1ef280703

Observation 35564da3-153b-4a32-9e5b-423c364ba9ea · outbound

This paper cites FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:41.029692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:41.029692Z digest=sha256:67991a099d7213cba4aa0d9d1b635e78cd2e0762a6df7cb191b7efbebd37c3c1

Observation 1223efc6-f851-4b4d-826c-94df6c0cd338 · outbound

This paper cites Plunkett, T.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Plunkett, T

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.873649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.033847Z digest=sha256:46f1313c5e1cbf608e345c6f9ef391b3a8d8b07e201efbf3870f51d891307ab3

Observation 5d5f1297-6aff-4c50-9f7b-14cba7969d0e · outbound

This paper cites Plunkett, T.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Plunkett, T

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.861110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.037790Z digest=sha256:5e2033130fc80f40ba7b1c8a9a1982127b1efed92b5c8e73d15491f4fe1d3d4a

Observation dcf8159e-ec49-4a2d-b95f-adb79cbab499 · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:44:41.847818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.041602Z digest=sha256:5acfca7e3d9763d1ab6ac818e3e3d2f8661aaaf050db18bcec25cdffd0cf6e1e

Observation 115790db-21f6-4c9e-9380-a55bcd7c4eda · outbound

This paper cites Paulheim, Knowledge graph refinement: A survey of approaches and evaluation methods, Semantic web 8 (2017) 489–508.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Paulheim, Knowledge graph refinement: A survey of approaches and evaluation methods, Semantic web 8 (2017) 489–508

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.835129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.045373Z digest=sha256:3608a3514df184e1c8fa1a56d4b7fd5406040890d9f90685d504f8b104a61fb0

Observation aea1492c-2fcd-421b-8348-d4456da17f11 · outbound

This paper cites Conceptual Engineering Using Large Language Models.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Conceptual Engineering Using Large Language Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:44:41.407098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.049154Z digest=sha256:f8969f6feee63579c5d8e4a6efa0a3efdc8a147b632b18d100731b23fe6602dc

Observation a16f1638-79ec-408c-a502-9015ee746ee2 · outbound

This paper cites Khatri, C.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Khatri, C

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.821559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.053192Z digest=sha256:094a497d604cb52ba30b0b2955186820a2fff5945deb7da27c8ca079284aa8a4

Observation 7754fa78-ed36-4b53-8b1c-bfef38c369b9 · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:44:41.806745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.056990Z digest=sha256:0edddc507ecd4b5602091beb7831f0ac5e4dda8284e9b0cc275b287811f4ca1d

Observation 58b07540-9f08-49df-95c3-7957f0dc2b55 · outbound

This paper cites FAIR 2.0: Extending the FAIR Guiding Principles to Address Semantic Interoperability.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs FAIR 2.0: Extending the FAIR Guiding Principles to Address Semantic Interoperability

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:44:41.382356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.061584Z digest=sha256:b6cf20e76b4a12c9d2f33b347223b3150bf5135c3d92a6a125a06cb32886d9c8

Observation 10b97c26-7a43-4478-a969-70c262d4757b · outbound

This paper cites Elsahar, P.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Elsahar, P

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.790628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.066069Z digest=sha256:c0bd984103671bfce54bba46b58b258acb907a6f77d1361e1870b8ef826173ae

Observation 2899adbd-6ce3-4629-b788-84263596e522 · outbound

This paper cites Kojima, S.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Kojima, S

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.773787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.070405Z digest=sha256:8fdfe7ad6edbd4f3803fbfea7854019f9e934814b41f6b463e8d9fec2d926e3d

Observation c0a761df-ce65-4946-936d-7d4d60a26015 · outbound

This paper cites Evaluating Class Membership Relations in Knowledge Graphs using Large Language Models.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Evaluating Class Membership Relations in Knowledge Graphs using Large Language Models

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:44:41.358514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.074731Z digest=sha256:804c5375bb76fe046997f4a84c44be95c759670db73ad2ab434255c04b6a9718

Observation 2044898d-6f81-4aa9-bd60-dcbe72f24209 · outbound

This paper cites Can Large Language Models Be an Alternative to Human Evaluations?.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Can Large Language Models Be an Alternative to Human Evaluations?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:41.079145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:41.079145Z digest=sha256:3099a0c8a768576a36e9c4ba00dda924f53fdbc3d04060214f96506ab453a7e4

Observation 2c92bdb5-2795-40cf-8fef-06fc07aecfda · outbound

This paper cites LLMs instead of Human Judges? A Large Scale Empirical Study across 20 NLP Evaluation Tasks.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs LLMs instead of Human Judges? A Large Scale Empirical Study across 20 NLP Evaluation Tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:41.083648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:41.083648Z digest=sha256:662952adc30de33ec51f14ab2d325680d141c7a523d0684182abe900bcef39e0

Observation 880d571c-3ab7-4bc6-8fb6-19a2c5ab27e3 · outbound

This paper cites Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:41.088451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:41.088451Z digest=sha256:d27d8786dbecf8481c7ea5f4ca5ce876d6d9ccc46c280dbf1c889cd3a6ca5190

Observation 3e4b5eb8-cd73-499a-91e0-a84eaaccf2c7 · outbound

This paper cites SHROOM-INDElab at SemEval-2024 Task 6: Zero- and Few-Shot LLM-Based Classification for Hallucination Detection.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs SHROOM-INDElab at SemEval-2024 Task 6: Zero- and Few-Shot LLM-Based Classification for Hallucination Detection

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:44:41.188893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.093526Z digest=sha256:735b4996a6991b906f17a70022f6746538c8c381c0162fe13f29eb90c5dd7bb7

Observation fadf6e18-da0e-4dd9-8aa3-4b714d7a6fd0 · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:44:41.757669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.099593Z digest=sha256:79de951d54a6ce00d7c1100d33064bc722d21d1cad5993f77e879189e8825fc8

Observation e3c0eb50-c8d2-4773-b658-300f26ff2440 · outbound

This paper cites Do Language Models' Words Refer?.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Do Language Models' Words Refer?

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:44:41.288932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.104145Z digest=sha256:1b9461b3af8decdefa1c0e737f8fa00f729ecd9cf290b178c70d901a5c860613

Observation 8c5be06f-a1a6-461b-b91b-d31a612b9d8b · outbound

This paper cites Are Language Models More Like Libraries or Like Librarians? Bibliotechnism, the Novel Reference Problem, and the Attitudes of LLMs.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Are Language Models More Like Libraries or Like Librarians? Bibliotechnism, the Novel Reference Problem, and the Attitudes of LLMs

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:41.109110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:41.109110Z digest=sha256:0662ff4a350a58e2b182ae4b58d5914fbadd79c6263a328bc83de17a152b7070

Observation cd10a721-29ea-475e-a296-577cbf42e9db · outbound

This paper cites an unresolved cited work.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:44:41.740636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.113762Z digest=sha256:daa36ebec01a2ee080c476f40f5c3cd9ed5a468b7451ad984770f968e6fc6b16

Observation 80dada28-09cd-4bed-991d-26d8474f11ab · outbound

This paper cites On the referential capacity of language models: An internalist rejoinder to Mandelkern & Linzen.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs On the referential capacity of language models: An internalist rejoinder to Mandelkern & Linzen

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-09T10:44:41.250029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.118138Z digest=sha256:e96e64292f1143f81a593801166271f6b018ad1a66d3aae05cb78a749940f6e5

Observation f2d0ac24-53e6-4e65-8631-e23c60b58723 · outbound

This paper cites Grindrod, Large language models and linguistic intentionality, Synthese 204 (2024) 71.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Grindrod, Large language models and linguistic intentionality, Synthese 204 (2024) 71

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.722490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.122847Z digest=sha256:10ea026e1ecd24acaacd3cc121d5e94ba88b8b9db29daa16c10472433342d5cc

Observation f76fc0fc-9a92-496f-a0ee-56ed4520bd15 · outbound

This paper cites Berto, Topics of thought: The logic of knowledge, belief, imagination, Oxford University Press, 2022.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Berto, Topics of thought: The logic of knowledge, belief, imagination, Oxford University Press, 2022

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.706031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.127253Z digest=sha256:43d744b72f051f025eaffb50a0805579d818d10dc08a805ab337c81f24ef65a4

Observation 58cd10ed-76c6-42cd-ba3a-aa838abebad8 · outbound

This paper cites Hawke, Theories of aboutness, Australasian Journal of Philosophy 96 (2018) 697–723.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Hawke, Theories of aboutness, Australasian Journal of Philosophy 96 (2018) 697–723

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.689149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.131619Z digest=sha256:bbc3a2107388a42903fefa4d3ad38b18153577786f3540454e642d4751a6b89b

Observation 4e098e24-39a8-407e-9f3d-0918b3d110ca · outbound

This paper cites Hawke, L.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Hawke, L

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.671857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.136271Z digest=sha256:8efd1bd2d5a159207a2dbd0c29bb645e665ad3a55e193bbabd7b3fcd64f2d0e8

Observation a7064b6a-cf26-48ef-bd1f-3039b5af2f78 · outbound

This paper cites Standards for Belief Representations in LLMs.

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Standards for Belief Representations in LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T10:44:41.140689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:44:41.140689Z digest=sha256:6395def119612c7f2a9d97ff8443b4dd4858a6ff4de323074ba9139a1c8ccd52

Observation 5782b04b-2898-4993-941d-ee1a9e4b7846 · outbound

This paper cites Harding, Operationalising representation in natural language processing, The British Journal for the Philosophy of Science (2023).

A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs Harding, Operationalising representation in natural language processing, The British Journal for the Philosophy of Science (2023)

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:44:41.654092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T10:44:41.145054Z digest=sha256:42285193284a9e8f51fc17a54a2eaeabb67590cc4620972e13a81206657e784e

Pith citing papers

No inbound Pith citation observations are available.