Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T12:15:05.313014Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2502.04349.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T12:15:05.313014Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T14:43:16.241588Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T14:43:16.748149Z
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f3d79fd6-07d1-4ed1-ba0a-6dc6fd0c52f0 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Beyond Prompts: Dynamic Conversational Benchmarking of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46aa19f4-babb-48ad-8329-088b7f2e7d6a · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Documenting Large Webtext Corpora: A Case Study on the Colossal Clean Crawled Corpus
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0de8e0f8-411f-44ab-8d66-ee52e1785173 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Making Pre-trained Language Models Better Few-shot Learners
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44c35fb5-b18c-4166-9b13-e5f582701fdd · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b346bbd0-3f7f-4506-adca-1e69265dabf0 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 40472905-a9fa-40a1-86a5-d8872f4790a1 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Unveiling LLM Evaluation Focused on Metrics: Challenges and Solutions
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21c9d41a-c017-4551-9245-53e35e6eb829 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Challenges and Applications of Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 289baa1b-57dd-4d0f-b70a-abf23f01faa7 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, Yaser Al-Onaizan, Mohit Bansal, and Yun-Nung Chen (Eds.)
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59bd7b06-6dc2-40d8-b27a-b60826f6dcdf · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Mind the Gap: Assessing Temporal Generalization in Neural Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f84393a5-0c02-40ba-9418-dedb87497952 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a9d1dcf6-68e8-4fb4-9123-3b82a2aeb128 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture GPT-4o System Card
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05360c84-1bbe-48b5-80f3-edc63783cf52 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Generative Agents: Interactive Simulacra of Human Behavior
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e623876-1f18-41f0-a666-c5a59816843e · outbound
Dynamic benchmarking framework for LLM-based conversational data capture CoQA: A Conversational Question Answering Challenge
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6be2bb4f-c70a-4819-a5ac-86be86f64291 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Measuring Semantic Coherence of a Conversation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 492fdea4-bdbb-4a8f-ae12-473b1ae492b1 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e1d9affa-e025-4fc4-b69e-e0e77831ee06 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture MM-LLMs: Recent Advances in MultiModal Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c25b8169-3d8b-46b9-8fb2-a2c4448a6d66 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture O’Brien, Carrie J
Reference 311
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0e5577cb-cea9-4079-8fcc-20d986778bfb · outbound
Dynamic benchmarking framework for LLM-based conversational data capture 2016), Dynamic benchmarking framework for LLM-based conversational data capture 463–471
Reference 2016
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1adef826-c6b8-403a-b8b2-a6bd28e12e0e · outbound
Dynamic benchmarking framework for LLM-based conversational data capture InDesign, User Experience, and Usability: The- ory and Practice, Aaron Marcus and Wentao Wang (Eds.)
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2062522c-b1e0-44fc-89c4-0e1fe47183b9 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture InProceedings of the 2019 Conference of the North
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ecb8f8d7-c038-4436-8854-55dddb595b61 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Evaluating Large Language Models Trained on Code
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7fd95b5-b416-45a6-8129-07d2b8adb15f · outbound
Dynamic benchmarking framework for LLM-based conversational data capture On the Opportunities and Risks of Foundation Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ebbb4c2-02b4-4766-b7bb-fddb2c6c9ff2 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Clembench: Using Game Play to Evaluate Chat-Optimized Language Models as Conversational Agents
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf4d25b3-a630-4b69-af9d-0d1c66112866 · outbound
Dynamic benchmarking framework for LLM-based conversational data capture Large Language Model in Financial Regulatory Interpretation
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 46c991f9-9705-47d4-bc29-d02ccb290896 · inbound
ProactiveEval: A Unified Evaluation Framework for Proactive Dialogue Agents Dynamic benchmarking framework for LLM-based conversational data capture
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.