Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:27:14.740756Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:2608.11047.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:27:14.740756Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a82a223f-ce73-4e1a-86c8-524e11e4605b · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Royal Society Open Science , author =
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 18c20d07-9891-43d6-8914-389ee9b160c2 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d3d4b963-4817-483b-ad8c-f8a3a56859bd · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Journal of Financial Data Science , author =
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 21207e7a-cb1e-403f-a346-6a4f00d03b27 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1be2fae3-78d6-474f-b0d8-ca1693951683 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark and Schärli, Nathanael and Zhou, Denny , month = jul, year =
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8a2e1531-9206-4fea-8023-4fd74050eb7a · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Proceedings of the 2024
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11901028-0c03-4170-a2d9-03d864c444ab · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Proceedings of the 61st
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8275cf26-577b-4e5c-90f2-ee5a54ca28ed · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 638d1d04-55ae-4633-9f53-0766c56e991e · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Least-to-Most Prompting Enables Complex Reasoning in Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6070be9a-045b-4b3d-b50c-a1f8e4f02ad1 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark arXiv.org , author =
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 295d2c15-6841-4434-89b4-817a26a17ff7 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d68f7240-e2a6-436b-881e-57c7c7fcb871 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dc34919-0876-49ad-8c58-b9e45926ec96 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark State of
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1df4501-1674-4086-84be-0c3ad29511a4 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Scaling Down to Scale Up: A Guide to Parameter-Efficient Fine-Tuning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e070690-7432-4f89-9411-eadca9a51fac · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a7d4f31-e938-485a-b948-ddb02759a5a4 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Parameter-Efficient Transfer Learning for NLP
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fa46fb9-7413-44e4-8bdf-539da2ea1bae · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark LoRA: Low-Rank Adaptation of Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8aa8649-599e-4745-9a40-a696840490cf · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Faith and Fate: Limits of Transformers on Compositionality
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cfed1da-fdef-44d5-9985-0b5c4f905fae · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Compositional
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd3966e0-d8c0-4322-870a-048d2854dae9 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Findings of the
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91bf9fc0-82d3-427b-aa8d-ebad0bdf38b9 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24d6dea3-d680-4ed2-ad79-01097e3172ee · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark BizBench: A Quantitative Reasoning Benchmark for Business and Finance
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f11c845-30cd-4694-ad99-33752ee8ad1e · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark MultiHiertt: Numerical Reasoning over Multi Hierarchical Tabular and Textual Data
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce0b6129-7434-4486-baee-307b062af219 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8e4b936-32ed-49d7-8968-d03e5810c415 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Findings of the
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c2a1a7-ee44-4d37-8fbb-5b987b6af374 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark TabFact: A Large-scale Dataset for Table-based Fact Verification
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1adfcf5-f805-47ce-beca-cfd2b880a6bb · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark doi:10.48550/arXiv.2509.25160 , abstract =
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e3a3cd35-f541-4820-a96e-53553c5be3c1 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42a7500a-9b27-4ba5-b9bb-d96447bfe60d · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark TAT-QA: A Question Answering Benchmark on a Hybrid of Tabular and Textual Content in Finance
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8500b8c7-4524-488f-aecf-0c407067f575 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark DORO: Distributional and Outlier Robust Optimization
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 17ebcb4b-0bde-442b-839a-eb79b1d79672 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark and Haghgoo, Behzad and Chen, Annie S
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 893c57f4-7fcf-481f-8125-1d612ac6cda5 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5b2f6bc-9153-46e8-b5f2-fa4d890b401a · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Evaluating
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a11c748-9428-41a6-be10-d3b0609dc4f5 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Enabling
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70fd02e5-b0f8-4348-a893-5c06c9662ec1 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Measuring
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 435d29ca-81d3-40cf-8d41-72191df3d903 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Advances in Neural Information Processing Systems , author =
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ae0a78c-693f-4c9c-a4ba-6c5ab9043d4e · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark DeepSciVerify: Verifying Scientific Claim--Citation Alignment via LLM-Driven Evidence Escalation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0786b739-323b-406d-81cf-1a8b92555a4d · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Evaluating Large Language Models at Evaluating Instruction Following
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04ce2feb-6030-4475-83fc-3616d40e9a35 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 372cb7e0-17b5-4fbf-9a42-f20d3b8c57c3 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 28a4c5a3-e92b-42c4-953e-6024c09ca0c6 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Proceedings of the 63rd
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e170810-b0cb-4fdf-893b-cedca4f5aa9b · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a744e5b3-94de-4f9c-abd5-b48e66509e8b · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark On the Limits of LLM-as-Judge for Scientific Novelty Assessment
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25cb44d5-be12-404e-bdc6-331dd3e08d6b · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark and Goel, Pranav and Stoehr, Niklas and Ash, Elliott and Hoyle, Alexander Miserlis , editor =
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c4fff4a4-17d5-4146-894e-98057756c811 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9743ab4e-951d-4903-8156-94fb349244f4 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark AURA: Adaptive Uncertainty-aware Refinement for LLM-as-a-Judge Auditing
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c487bd08-d28b-4bf6-966a-6ce5fbfbb600 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Rethinking Literature Search Evaluation: Deep Research Helps, and Human Citation Lists Are Not a Ground Truth
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 71e4d072-acbd-4332-aee2-764f35a263fd · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Citation Grounding Measures the Oracle: Graph Coverage Determines Reported LLM Hallucination Rates in Law
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a5c890be-981e-49d4-b186-c5aec3ffd088 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark LegalBench: A Collaboratively Built Benchmark for Measuring Legal Reasoning in Large Language Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f21bfed-a74f-4d0c-bbe5-4c08ef2704e8 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Findings of the
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b9d8cbf2-89a2-4078-8134-c665b339f574 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark FinGPT: Instruction Tuning Benchmark for Open-Source Large Language Models in Financial Datasets
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e500578-578c-4cfb-892c-0023d18ce1c8 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark FinanceBench: A New Benchmark for Financial Question Answering
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df4e0999-fe47-4441-a68f-e71cba2bccb3 · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark Beyond Knowledge to Agency: Evaluating Expertise, Autonomy, and Integrity in Finance with CNFinBench
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 788d2f23-0e55-4f54-9dbb-b24f9e08b1af · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark ConvFinQA: Exploring the Chain of Numerical Reasoning in Conversational Finance Question Answering
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e72da66-2209-4f13-bfee-8d4b5c18002b · outbound
V-FiLLM: Verified Financial LLM Reasoning Benchmark FinQA: A Dataset of Numerical Reasoning over Financial Data
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.