Pith. sign in

Paper Citation Record · LEDGER

HalluLens: LLM Hallucination Benchmark

As of 16 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 26 inbound Pith citation observations for arXiv:2504.17550.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17550 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:40:19.484778Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:58:22.465675Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation fe2f6aee-274e-4910-81f7-293bbd59e291 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:20.127487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.369975Z digest=sha256:ea3698fb7de27e6b1a68bf689857bdd2f85492734a0516d66099d581d2805fad

Observation ec1df31a-8bdc-4d99-ad01-b4a716cdcecc · outbound

This paper cites Survey on Factuality in Large Language Models: Knowledge, Retrieval and Domain-Specificity.

HalluLens: LLM Hallucination Benchmark Survey on Factuality in Large Language Models: Knowledge, Retrieval and Domain-Specificity

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T10:40:19.329519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:40:19.329519Z digest=sha256:82a5d0a59a1f278aa0366a069f2b8f0e7ab5b12ad82f7b5ceb019f18715cc133

Observation 38827227-be20-4e5b-89e4-87938905063b · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:19.989618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.400936Z digest=sha256:7f387b91dcd2ad60072f7e6055a98ce6303aa4039b97c94688e711dbffff9e15

Observation e3ba799f-b172-4b43-bb5d-76a7ff6d3be4 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:19.948296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.422585Z digest=sha256:2b6d5c795f0ff21212d168450e4c3c82142dbe88b1dd15c0d012377eefed55af

Observation 48139ecb-1f90-44c1-9885-59e738babfcf · outbound

This paper cites unanswerable.

HalluLens: LLM Hallucination Benchmark unanswerable

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:40:19.919710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.430376Z digest=sha256:86e2b9f5abdbbb8787cc34181ab214c0440ff27864fa89df9c80cf4c15bb1118

Observation 238c2af3-5710-4658-981a-794abdbe76d0 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:19.864568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.440500Z digest=sha256:c53295ed3b878daad4f1b8d81d741c620e18a05b4881bf3b1e8fa2ba1528497f

Observation 50c546e1-9c8d-493c-8169-255c6e092061 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:20.081379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.450485Z digest=sha256:353181552e10f6f12694df2d9ffb2f4c49ae3077402811bb60899ed9ff48f915

Observation b417c244-28f7-4c69-802a-e54dbf6aeedf · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:20.039403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.466802Z digest=sha256:11845445cb212a8378dbbc577758eca109748ab77a6901f5c2ea16f8a81964d6

Observation 467afffd-343b-4004-8875-2038360921d3 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:19.822478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.474582Z digest=sha256:b1e3221367b504cd824107e2016788e231c3f1f750ab7023e4779c3fa91bbeef

Observation 35ce6d55-aa47-453e-a4f6-ba6bb0068307 · outbound

This paper cites {wiki_document}.

HalluLens: LLM Hallucination Benchmark {wiki_document}

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:40:19.782703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.484778Z digest=sha256:b731ec33d0c85c3aad6c905d1d19103869e2e9344cf06759b2200aa78f1888b0

Observation 553fc54c-7141-4e2d-a89b-d175d90ffc60 · outbound

This paper cites GPT-4 Technical Report.

HalluLens: LLM Hallucination Benchmark GPT-4 Technical Report

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-16T10:40:19.316874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:40:19.316874Z digest=sha256:50bfc70ee99500396f5a3082f67854646a9fd033ac013e422bbdfe1ea7b2d196

Observation ca7a9cbc-c633-41bf-8faf-dbd2ad02ca5e · outbound

This paper cites Measuring short-form factuality in large language models.

HalluLens: LLM Hallucination Benchmark Measuring short-form factuality in large language models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-16T10:40:19.340665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:40:19.340665Z digest=sha256:51199ca401e89f7a98d81bae5c3b251c709227c4ccc2d59dc9ce38b8d7cdc904

Observation 7bca60dd-0ac8-44d7-8a17-bed5a53ad2dd · outbound

This paper cites doi: 10.1145/3571730.http://dx.doi.org/10.1145/3571730.

HalluLens: LLM Hallucination Benchmark doi: 10.1145/3571730.http://dx.doi.org/10.1145/3571730

Reference 2023

Resolution
verified exact
doi, observed 2026-08-16T10:40:19.598496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.305831Z digest=sha256:c3f97c75c46093a6e026570b30e56775c037c977e3cc65ea5267b60d10c0f75e

Observation 173c253e-88c1-4787-bd80-78b20004cf94 · outbound

This paper cites correct”, “incorrect.

HalluLens: LLM Hallucination Benchmark correct”, “incorrect

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:40:20.159736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:40:19.358216Z digest=sha256:6b605661618869a5be629882f1b38f77e690e0453616b2d66a6ec6e4c2630b22

Pith citing papers

Observation 37fdbe5b-af66-4104-896a-fb76efdacb38 · inbound

Phare: A Safety Probe for Large Language Models cites this paper.

Phare: A Safety Probe for Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.465675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.465675Z digest=sha256:60a3196a65cd6c9fdaa698b9087bca97c5279da61b927f201aa630995f01f753

Observation 6c6d2e74-fa0b-4a8b-bb16-c419cb0eefa0 · inbound

The Hallucination Tax of Reinforcement Finetuning cites this paper.

The Hallucination Tax of Reinforcement Finetuning HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:44:30.304843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:44:30.304843Z digest=sha256:9b1a024f179bdadd7d15a3fc9c86f8fdd28de13762ad52f5d79a8e3c5cb39d5d

Observation 710392c0-8593-4199-9562-ece45783b8a7 · inbound

AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions cites this paper.

AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions HalluLens: LLM Hallucination Benchmark

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:08.125965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:08.125965Z digest=sha256:161755e859d81c2ae8f55b9e8e330cbedd8893d6966a15840e69eee39826d262

Observation 77e35e4d-2d50-473b-92a1-f8497ba24f36 · inbound

Machine Mirages: Defining the Undefined cites this paper.

Machine Mirages: Defining the Undefined HalluLens: LLM Hallucination Benchmark

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:24:29.959566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:24:29.959566Z digest=sha256:256cc8db654370cc7fe3b695c3f603b0d809def944a4062678f7dfbc186fa22f

Observation dd8252b2-fc3d-4546-81ec-605e437dbb68 · inbound

Embodied AI Agents: Modeling the World cites this paper.

Embodied AI Agents: Modeling the World HalluLens: LLM Hallucination Benchmark

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T22:10:25.436137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:10:25.436137Z digest=sha256:a35dc714c9ca719bf8ea0b35f349ef194592ee79e00375d14e7df1ec12e4c9db

Observation 37e98791-aba2-4f3b-9a78-ecf26545fd0c · inbound

Introducing the Swiss Food Knowledge Graph: AI for Context-Aware Nutrition Recommendation cites this paper.

Introducing the Swiss Food Knowledge Graph: AI for Context-Aware Nutrition Recommendation HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:43:11.362519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:43:11.362519Z digest=sha256:c8d8664c1c97d4a2240e2fd87b18c0cec020fae588b8f080bb0d3a40f06d45a8

Observation c3bc9cd2-4707-482c-8568-42007d7edf06 · inbound

MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them cites this paper.

MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them HalluLens: LLM Hallucination Benchmark

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:36.599018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:36.599018Z digest=sha256:932e5b1c73cce8de68d0df24d1143ef7209baab6b0bd00998609274145b3c3de

Observation 5831c5ed-3040-4a23-a8c2-3a38c4a34cce · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:09.780706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:09.780706Z digest=sha256:a77d6f79b95e5c6c4f701b0cab231423d476cf4163b5b7e157f00232c4ad6640

Observation 08f453e8-d49c-4bdc-8a33-614d4f0addf4 · inbound

ReasoningTrack: Chain-of-Thought Reasoning for Long-term Vision-Language Tracking cites this paper.

ReasoningTrack: Chain-of-Thought Reasoning for Long-term Vision-Language Tracking HalluLens: LLM Hallucination Benchmark

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T23:33:19.965610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:33:19.965610Z digest=sha256:c0c7adfccc651e52cdf41d1f5a00c344f921d89b7735bf4464d748f7bc6537db

Observation 1caab12d-ef18-4070-a689-d66bcbeb9b4a · inbound

GOSU: Retrieval-Augmented Generation with Global-Level Optimized Semantic Unit-Centric Framework cites this paper.

GOSU: Retrieval-Augmented Generation with Global-Level Optimized Semantic Unit-Centric Framework HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T13:36:35.132035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:36:35.132035Z digest=sha256:fb3c6d6af3db0cb0c07992d52df8c579116caacd4f59b36c82a105314cf45b22

Observation 1e25fd92-76c3-4f24-9d4d-bd7323c38648 · inbound

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations cites this paper.

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.497247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-18T12:54:01.717015Z digest=sha256:090ab2e9dc6f983dcae5295a9a86af5bdb4cdcbcec08ca22a7f83446de8faf9f

Observation eb0f60c0-8233-4d36-b187-45b6acb266dc · inbound

When Do Hallucinations Arise? A Graph Perspective on the Evolution of Path Reuse and Path Compression cites this paper.

When Do Hallucinations Arise? A Graph Perspective on the Evolution of Path Reuse and Path Compression HalluLens: LLM Hallucination Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T13:07:00.928770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:07:00.928770Z digest=sha256:cdfa1fea2ad68c136a61b16e8f2f2a8843de8974091788865e3839d7d3b54015

Observation 5cb32dba-21c3-4776-b09b-ea6890e3981c · inbound

From Binary Groundedness to Support Relations: Towards a Reader-Centred Taxonomy for Comprehension of AI Output cites this paper.

From Binary Groundedness to Support Relations: Towards a Reader-Centred Taxonomy for Comprehension of AI Output HalluLens: LLM Hallucination Benchmark

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:30:58.657451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T17:37:28.129376Z digest=sha256:7a3fabe11ffce318960e38af59f0f451b2d0a0358260af6a58c4658831331460

Observation 6be78a1e-743e-44e0-8a4b-e759714bc979 · inbound

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents cites this paper.

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents HalluLens: LLM Hallucination Benchmark

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:31:30.913700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T06:27:19.717895Z digest=sha256:53c6f620ef3d019dcb7df86b413e22a3e3181dd111562034c982c3baaa7c145c

Observation fde6b7d0-6551-476f-9e37-e6c3b7aa2dd7 · inbound

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks cites this paper.

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks HalluLens: LLM Hallucination Benchmark

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:38:42.843693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T05:12:31.218055Z digest=sha256:510467746e4c6ec7f44d8de818a10890a80b27349d9344a5f60a7360245a1bba

Observation 58b08582-fad0-494c-bc2f-8345ec78c00c · inbound

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation cites this paper.

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation HalluLens: LLM Hallucination Benchmark

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:56:34.399829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T03:42:06.876417Z digest=sha256:2b22d6d90dfd920b534b77bd640a5c1c4aaa3f1da3a44064387095428868397e

Observation 6428576b-9a66-41e8-92e0-ce1a174edf42 · inbound

Hallucination as an Anomaly: Dynamic Intervention via Probabilistic Circuits cites this paper.

Hallucination as an Anomaly: Dynamic Intervention via Probabilistic Circuits HalluLens: LLM Hallucination Benchmark

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:51:11.274747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T10:52:14.666333Z digest=sha256:3d2b29b738a1792da92104248cde9b0d215b53de89d9246806dcd4efb8b4d7a2

Observation b96bea36-ddba-4ac7-a459-694aa94f0410 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs HalluLens: LLM Hallucination Benchmark

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.293122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-12T04:50:17.399580Z digest=sha256:05e958549de02f4bc227d11068e85c0bfccd6091443e183cbf89f6c18bb0c387

Observation ce33085a-6e5d-4c06-9941-2928e6d6c9d7 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs HalluLens: LLM Hallucination Benchmark

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:22:28.981194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-13T07:20:32.494840Z digest=sha256:cda93559a3890494d8cae8330fab778d0a8a8b29ae2e94954d82d0a9faa4594b

Observation bd559f56-718a-48b3-aefb-91e6c6cd3c4c · inbound

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations cites this paper.

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations HalluLens: LLM Hallucination Benchmark

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:56.743792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-14T20:13:10.814899Z digest=sha256:5805fccec6fdc99e57b0d61e5d95b4c8a35d86304033f39c33d8b63081a71232

Observation a6760fd2-3410-4673-83a9-401894c769b2 · inbound

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models cites this paper.

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:53:59.416150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-21T05:49:55.738870Z digest=sha256:c84dec65a12d62739dc8b3307880d6254b07fa9a733191fc0da0c63d6ea1f392

Observation ccf7b7c9-e403-42c9-b9d7-0bdaec1b7542 · inbound

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance cites this paper.

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance HalluLens: LLM Hallucination Benchmark

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T09:03:16.148279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-29T08:54:48.807164Z digest=sha256:b1a97dea35ed46d6fa50fda41f2397943fcd9742fd99acc8bfb8783be4a1de5b

Observation b82edf81-756c-4f85-939b-097951a4d9e7 · inbound

Latent Performance Profiling of Large Language Models cites this paper.

Latent Performance Profiling of Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.394631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-29T07:31:02.595386Z digest=sha256:59a4bbd164c45a0885d885dc672af504b20237dbe5a47b81751d8d33f9c99e3f

Observation f3365005-c053-4dbb-9a77-89f8f2fcb14d · inbound

Disentangling Visual and Factual Correctness in LVLMs' Visualization Literacy cites this paper.

Disentangling Visual and Factual Correctness in LVLMs' Visualization Literacy HalluLens: LLM Hallucination Benchmark

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:06:26.198841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-28T11:19:51.893740Z digest=sha256:de6cae5986809aa80bf29887b37c646e50e1d9edbb3ab893e100fe369121d3a5

Observation 8026be1b-17af-41bf-9403-7f7421cd59c0 · inbound

What Do People Actually Want From AI? Mapping Preference Plurality cites this paper.

What Do People Actually Want From AI? Mapping Preference Plurality HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-28T01:41:29.860815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-28T01:32:04.660400Z digest=sha256:7499d5b84cda072411bdec8c262f3c0272f0dd939159d99dff098176b6756bbc

Observation 00ed01a9-0708-42ef-9e46-ad803e5d813a · inbound

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions cites this paper.

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions HalluLens: LLM Hallucination Benchmark

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T01:57:59.043987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:57:59.043987Z digest=sha256:8a9ad8dfcec8f4463d5cdba3422d81cfdf01780ef0033c9bc2ef180680732dce