Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:13:17.797667Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 8 inbound Pith citation observations for arXiv:2506.11440.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:13:17.797667Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:22:44.774647Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T13:19:50.245560Z
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9bd0fe80-df29-4d2b-a589-d3c6b0e632c2 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Many-Shot In-Context Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd26a989-0335-450d-9d89-aa1080f9a8d6 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Yi: Open Foundation Models by 01.AI
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56e6fd7a-77b1-4878-bb0d-11a3c679c1f8 · outbound
AbsenceBench: Language Models Can't Tell What's Missing L-Eval: Instituting Standardized Evaluation for Long Context Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdcc8374-c029-4222-ba5e-64aaaac30c8c · outbound
AbsenceBench: Language Models Can't Tell What's Missing Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 59e6c12b-1be1-44d5-aa28-03d6df593380 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Claude 3 haiku: our fastest model yet, 2024
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4b240ed1-6152-4706-af7c-a8f15e1ee392 · outbound
AbsenceBench: Language Models Can't Tell What's Missing LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26baed68-fc15-4741-9621-68208e1e5924 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Unlimiformer: Long-Range Transformers with Unlimited Length Input
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b11f4f4b-22c4-454e-8ca2-8d6c6f6f1730 · outbound
AbsenceBench: Language Models Can't Tell What's Missing BooookScore: A systematic exploration of book-length summarization in the era of LLMs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6acd2405-5ee4-40df-ae14-56d6fbbad1a0 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Extending Context Window of Large Language Models via Positional Interpolation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4aeb2d8-1861-4e57-b56d-81de0cdde438 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86d5eac7-5bc7-4027-b66b-cf8d190b1f3c · outbound
AbsenceBench: Language Models Can't Tell What's Missing DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce77bd0a-7837-4a4f-9fcd-ff3e30cd340a · outbound
AbsenceBench: Language Models Can't Tell What's Missing Mathematical Capabilities of ChatGPT
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c8f867e-c763-44a5-b07b-fafc618db48c · outbound
AbsenceBench: Language Models Can't Tell What's Missing Simple Hardware-Efficient Long Convolutions for Sequence Modeling
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91348c31-eb1e-4342-aa92-49a2c48aff3f · outbound
AbsenceBench: Language Models Can't Tell What's Missing How to train long-context language models (effectively), 2025
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 314cd94a-af43-44cc-b755-1474571bdac2 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Gemini 2.5: Our most intelligent ai model, 2025
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3a79523c-c853-49ec-a913-ac6ebc8bd7ff · outbound
AbsenceBench: Language Models Can't Tell What's Missing The Llama 3 Herd of Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95615ecf-e4c3-4f3f-afe8-9f3d94364ce2 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 534072d9-ab58-442f-a018-96860604436e · outbound
AbsenceBench: Language Models Can't Tell What's Missing RULER: What's the Real Context Size of Your Long-Context Language Models?
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23c9b130-cb9c-4e0c-97d7-b9063af6f734 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Mixtral of Experts
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 984be04b-3503-41a5-b086-912982775004 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Needle in a haystack - pressure testing llms, 2023
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 23c904a3-feed-4302-b6b7-fc1b8b932fd3 · outbound
AbsenceBench: Language Models Can't Tell What's Missing FABLES: Evaluating faithfulness and content selection in book-length summarization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c134385-792b-4c51-9ea6-155b1800f4e3 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e5344b0-c219-46e2-99c1-86b488710e7e · outbound
AbsenceBench: Language Models Can't Tell What's Missing The NarrativeQA Reading Comprehension Challenge
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bffc137-3dd6-44fc-9766-b3b995853ebb · outbound
AbsenceBench: Language Models Can't Tell What's Missing Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f02fd82-05aa-4263-8a70-41a75902de14 · outbound
AbsenceBench: Language Models Can't Tell What's Missing The llama 4 herd: The beginning of a new era of natively multimodal ai innovation, 2025
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 48a5b14c-7bf1-4424-8f5f-cd326fc48a71 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Openai o3-mini, pushing the frontier of cost-effective reasoning., 2025
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1f0529d3-daf8-4b7f-8057-ad04af79bf06 · outbound
AbsenceBench: Language Models Can't Tell What's Missing GPT-4o System Card
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e74647e-5850-4944-af1d-86d0873be01a · outbound
AbsenceBench: Language Models Can't Tell What's Missing GPT-4 Technical Report
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ccfa6f8-8ac4-4fc3-a740-588bac6e9edd · outbound
AbsenceBench: Language Models Can't Tell What's Missing gutenberg-poetry-corpus: A corpus of poetry from project gutenberg, 2018
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0b8b784a-e963-4eea-88f9-63e0b67a47c8 · outbound
AbsenceBench: Language Models Can't Tell What's Missing RWKV: Reinventing RNNs for the Transformer Era
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0555fb9-7405-4359-b81d-a649a992860b · outbound
AbsenceBench: Language Models Can't Tell What's Missing Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fb76f57-1b76-41a0-8d9c-66565c113efe · outbound
AbsenceBench: Language Models Can't Tell What's Missing Qwen3 technical report, 2025 a
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 624fdfa1-57c7-4b31-aa60-6f6d80e6ebb9 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Qwq-32b: Embracing the power of reinforcement learning, 2025 b
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5ef853bc-f219-4f16-baf3-d60978a1bcf4 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Code Llama: Open Foundation Models for Code
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d8bb6cf-7172-4920-9c7b-ee88dc9a0926 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Z ero SCROLLS : A zero-shot benchmark for long text understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f5985b6-1d6a-4717-94f6-42bb50b8cfe6 · outbound
AbsenceBench: Language Models Can't Tell What's Missing RoFormer: Enhanced Transformer with Rotary Position Embedding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 779560ff-d125-4b76-a476-ab4d55c1d77c · outbound
AbsenceBench: Language Models Can't Tell What's Missing A Length-Extrapolatable Transformer
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef0c8129-be67-40d4-8a09-820e953193ad · outbound
AbsenceBench: Language Models Can't Tell What's Missing Attention is all you need
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41be1d9e-0019-4b67-bfaf-49151455a136 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Michelangelo: Long Context Evaluations Beyond Haystacks via Latent Structure Queries
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b28314-4984-4097-8f42-ffe392ba4c02 · outbound
AbsenceBench: Language Models Can't Tell What's Missing NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8c5b1f1-76a0-45ce-9e64-548cd4be8f3a · outbound
AbsenceBench: Language Models Can't Tell What's Missing Grok 3 beta — the age of reasoning agents, 2025
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 77426410-528d-43a0-a526-55cec043c59f · outbound
AbsenceBench: Language Models Can't Tell What's Missing Stress-Testing Long-Context Language Models with Lifelong ICL and Task Haystack
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0416608e-094e-4f4f-a4c2-4a20a02e5fc6 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Qwen2.5 Technical Report
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 102efd5b-655b-4ab6-ba39-f87b1ae4a122 · outbound
AbsenceBench: Language Models Can't Tell What's Missing Long-Context Language Modeling with Parallel Context Encoding
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23303ff9-36c3-4125-bd84-9493b8df1701 · outbound
AbsenceBench: Language Models Can't Tell What's Missing HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f936a41a-b2a4-4d44-b710-f6820d013140 · outbound
AbsenceBench: Language Models Can't Tell What's Missing $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e55bcf3-5a6b-4018-9736-12d0dcdedbaf · outbound
AbsenceBench: Language Models Can't Tell What's Missing Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a529bcd7-92a9-4507-b7d0-d067eb575c35 · inbound
Mitigating hallucinations and omissions in LLMs for invertible problems: An application to hardware logic design automation AbsenceBench: Language Models Can't Tell What's Missing
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d5bceb12-34fd-4946-9055-885ddeb8c53e · inbound
Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems AbsenceBench: Language Models Can't Tell What's Missing
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 27e3d2c1-b4be-484d-9a5b-c41848ca5ebf · inbound
Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems AbsenceBench: Language Models Can't Tell What's Missing
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 85b9209b-8f41-477c-a373-1c1bcf0298e8 · inbound
Locality Does Not Imply Reachability: Boundary Repair in Block-Sparse Causal Attention AbsenceBench: Language Models Can't Tell What's Missing
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c3b00df-813b-4824-b32f-191aab9e1cd9 · inbound
The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation de7b2486-b4fc-40ae-9dac-3c900777e07b · inbound
The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c04a9899-bad2-4526-b63f-f5d0e9e2010e · inbound
The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0cc3f5a-3a52-4705-8a0d-733ca1a87e7f · inbound
When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models AbsenceBench: Language Models Can't Tell What's Missing
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.