Pith. sign in

Paper Citation Record · LEDGER

AbsenceBench: Language Models Can't Tell What's Missing

As of 18 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 8 inbound Pith citation observations for arXiv:2506.11440.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11440 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:13:17.797667Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:22:44.774647Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:19:50.245560Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9bd0fe80-df29-4d2b-a589-d3c6b0e632c2 · outbound

This paper cites Many-Shot In-Context Learning.

AbsenceBench: Language Models Can't Tell What's Missing Many-Shot In-Context Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.648452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.648452Z digest=sha256:83f0fa6426e54a6e3d6f2102b9e9202a6dea715e111a8ef7698ba793f3817006

Observation dd26a989-0335-450d-9d89-aa1080f9a8d6 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

AbsenceBench: Language Models Can't Tell What's Missing Yi: Open Foundation Models by 01.AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.652986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.652986Z digest=sha256:166ddd982d2b88df3814a5f2884f2b9454eccb77ce7366d12080d2d10b1fd4c5

Observation 56e6fd7a-77b1-4878-bb0d-11a3c679c1f8 · outbound

This paper cites L-Eval: Instituting Standardized Evaluation for Long Context Language Models.

AbsenceBench: Language Models Can't Tell What's Missing L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.656859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.656859Z digest=sha256:6c02b81b40338e01ba04bd6840998651ad9d99847f4c88e5802a74b948039537

Observation cdcc8374-c029-4222-ba5e-64aaaac30c8c · outbound

This paper cites Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud.

AbsenceBench: Language Models Can't Tell What's Missing Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.265746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.660267Z digest=sha256:07de5265c3e75b0bf8abcf314b158c12e3f518cab6187dd7b5548cee09613112

Observation 59e6c12b-1be1-44d5-aa28-03d6df593380 · outbound

This paper cites Claude 3 haiku: our fastest model yet, 2024.

AbsenceBench: Language Models Can't Tell What's Missing Claude 3 haiku: our fastest model yet, 2024

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.256396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.663450Z digest=sha256:6d8bfa49a69eb8c0b1f98507b0636b03ca3c3dbcd7517ab6f0d4eb8da009b105

Observation 4b240ed1-6152-4706-af7c-a8f15e1ee392 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

AbsenceBench: Language Models Can't Tell What's Missing LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.666475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.666475Z digest=sha256:24c3b503421ace39048569778d201a5791af5fd02d3a7f8474a605201e8cc765

Observation 26baed68-fc15-4741-9621-68208e1e5924 · outbound

This paper cites Unlimiformer: Long-Range Transformers with Unlimited Length Input.

AbsenceBench: Language Models Can't Tell What's Missing Unlimiformer: Long-Range Transformers with Unlimited Length Input

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.670308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.670308Z digest=sha256:dd6fde11924360f49003bb495709e9ef0330932e401e4e3d4e152b8e8009c09a

Observation b11f4f4b-22c4-454e-8ca2-8d6c6f6f1730 · outbound

This paper cites BooookScore: A systematic exploration of book-length summarization in the era of LLMs.

AbsenceBench: Language Models Can't Tell What's Missing BooookScore: A systematic exploration of book-length summarization in the era of LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.673581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.673581Z digest=sha256:c10b91eb6239282ed0fced74a6bdf0050b0c2d94f436e2100e64e5fb3ab3b1f9

Observation 6acd2405-5ee4-40df-ae14-56d6fbbad1a0 · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

AbsenceBench: Language Models Can't Tell What's Missing Extending Context Window of Large Language Models via Positional Interpolation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.676871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.676871Z digest=sha256:f5fccd1c9fe7da40cbdf2f0d39c502727b16cd17d925cee3d1c5014904ba20a6

Observation b4aeb2d8-1861-4e57-b56d-81de0cdde438 · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

AbsenceBench: Language Models Can't Tell What's Missing Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.679986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.679986Z digest=sha256:da27d2c4c1fab4247964e4b3c4868c3a3b1c447f0eac158067867a294b0980e7

Observation 86d5eac7-5bc7-4027-b66b-cf8d190b1f3c · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AbsenceBench: Language Models Can't Tell What's Missing DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.683266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.683266Z digest=sha256:a2e219ee98e550a3f525c9f4f4490caac86cf933bf526748a086c1faf53301eb

Observation ce77bd0a-7837-4a4f-9fcd-ff3e30cd340a · outbound

This paper cites Mathematical Capabilities of ChatGPT.

AbsenceBench: Language Models Can't Tell What's Missing Mathematical Capabilities of ChatGPT

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.686818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.686818Z digest=sha256:c83323f68b175620bb88a83fabca4bc00dee288b692c6a0be3c1f85da6bebfe1

Observation 8c8f867e-c763-44a5-b07b-fafc618db48c · outbound

This paper cites Simple Hardware-Efficient Long Convolutions for Sequence Modeling.

AbsenceBench: Language Models Can't Tell What's Missing Simple Hardware-Efficient Long Convolutions for Sequence Modeling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.690031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.690031Z digest=sha256:34a8b805c0436e359872c777dce37ac9b03d5ca302e7bc017af0fc4bb35dcf0a

Observation 91348c31-eb1e-4342-aa92-49a2c48aff3f · outbound

This paper cites How to train long-context language models (effectively), 2025.

AbsenceBench: Language Models Can't Tell What's Missing How to train long-context language models (effectively), 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.693121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.693121Z digest=sha256:88bc04c0875e38e2f078615f126817f37d8a5291f5914e824c528add9179124e

Observation 314cd94a-af43-44cc-b755-1474571bdac2 · outbound

This paper cites Gemini 2.5: Our most intelligent ai model, 2025.

AbsenceBench: Language Models Can't Tell What's Missing Gemini 2.5: Our most intelligent ai model, 2025

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.247436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.695912Z digest=sha256:5e9efd5d1d96301f5736f2d65da9351b33161eaf8b7a2a13c86a2a207cd7317c

Observation 3a79523c-c853-49ec-a913-ac6ebc8bd7ff · outbound

This paper cites The Llama 3 Herd of Models.

AbsenceBench: Language Models Can't Tell What's Missing The Llama 3 Herd of Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.699059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.699059Z digest=sha256:c13ede65133a7fc8a68a97bd1dfd844853aeebe752d3b98f8dc7f31444ea6682

Observation 95615ecf-e4c3-4f3f-afe8-9f3d94364ce2 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

AbsenceBench: Language Models Can't Tell What's Missing Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.702868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.702868Z digest=sha256:affaba1af166a521603edbe7f8a1c0e2e79a650b73946002f3def7d138503251

Observation 534072d9-ab58-442f-a018-96860604436e · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

AbsenceBench: Language Models Can't Tell What's Missing RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.705952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.705952Z digest=sha256:fbbc69b0deb9ccf3d43ba3efca5745bae5969ebb14abf5a50e012a0e2b461d6b

Observation 23c9b130-cb9c-4e0c-97d7-b9063af6f734 · outbound

This paper cites Mixtral of Experts.

AbsenceBench: Language Models Can't Tell What's Missing Mixtral of Experts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.709044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.709044Z digest=sha256:5010393a538ed8354d545cb68f9a03250dea8587b849c917a61a10c78bfa8c23

Observation 984be04b-3503-41a5-b086-912982775004 · outbound

This paper cites Needle in a haystack - pressure testing llms, 2023.

AbsenceBench: Language Models Can't Tell What's Missing Needle in a haystack - pressure testing llms, 2023

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.238761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.712296Z digest=sha256:90927c9f28b6e5c738403f2e585cfdfa9056746cfc59c37e4f84bd7e251f2254

Observation 23c904a3-feed-4302-b6b7-fc1b8b932fd3 · outbound

This paper cites FABLES: Evaluating faithfulness and content selection in book-length summarization.

AbsenceBench: Language Models Can't Tell What's Missing FABLES: Evaluating faithfulness and content selection in book-length summarization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.715567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.715567Z digest=sha256:97295d926d606b1026c3e12bbbf65f253dce80b99d78390b5a502c5b53740d4a

Observation 6c134385-792b-4c51-9ea6-155b1800f4e3 · outbound

This paper cites Benchmarking Cognitive Biases in Large Language Models as Evaluators.

AbsenceBench: Language Models Can't Tell What's Missing Benchmarking Cognitive Biases in Large Language Models as Evaluators

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.718512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.718512Z digest=sha256:ec6b4847bc9a645d063885c61f72395857adfeb0c44695deb725ea5b70c614b2

Observation 3e5344b0-c219-46e2-99c1-86b488710e7e · outbound

This paper cites The NarrativeQA Reading Comprehension Challenge.

AbsenceBench: Language Models Can't Tell What's Missing The NarrativeQA Reading Comprehension Challenge

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.721714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.721714Z digest=sha256:5846eafd23e5e041477d9daa65f9cf875bc8d32f95e1cc043df55ecaadd1de10

Observation 7bffc137-3dd6-44fc-9766-b3b995853ebb · outbound

This paper cites Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?.

AbsenceBench: Language Models Can't Tell What's Missing Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.725291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.725291Z digest=sha256:42a6cc284493624124713e54dfaea9a1ddfadfba1731b54dd1ef2edba71ff96f

Observation 0f02fd82-05aa-4263-8a70-41a75902de14 · outbound

This paper cites The llama 4 herd: The beginning of a new era of natively multimodal ai innovation, 2025.

AbsenceBench: Language Models Can't Tell What's Missing The llama 4 herd: The beginning of a new era of natively multimodal ai innovation, 2025

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.228912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.728350Z digest=sha256:9f5ed4402b1edd52ffd2ec852d7ca6637a7cada0c48fd97ebe6f6472e625c85f

Observation 48a5b14c-7bf1-4424-8f5f-cd326fc48a71 · outbound

This paper cites Openai o3-mini, pushing the frontier of cost-effective reasoning., 2025.

AbsenceBench: Language Models Can't Tell What's Missing Openai o3-mini, pushing the frontier of cost-effective reasoning., 2025

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.219852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.731389Z digest=sha256:e41703829e047bfd62f5a0b302d483d1df294d295838309bc1d398181e9e7de5

Observation 1f0529d3-daf8-4b7f-8057-ad04af79bf06 · outbound

This paper cites GPT-4o System Card.

AbsenceBench: Language Models Can't Tell What's Missing GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.734425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.734425Z digest=sha256:b2e4779739a9a008e1e63b72c26db4027e22535789ac09b42499e8951daeb1b7

Observation 9e74647e-5850-4944-af1d-86d0873be01a · outbound

This paper cites GPT-4 Technical Report.

AbsenceBench: Language Models Can't Tell What's Missing GPT-4 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.737879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.737879Z digest=sha256:157aa6db039354d6fa90117a32009bb54c1f177f79eceaebe9f3b79e149f7829

Observation 9ccfa6f8-8ac4-4fc3-a740-588bac6e9edd · outbound

This paper cites gutenberg-poetry-corpus: A corpus of poetry from project gutenberg, 2018.

AbsenceBench: Language Models Can't Tell What's Missing gutenberg-poetry-corpus: A corpus of poetry from project gutenberg, 2018

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.210950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.741220Z digest=sha256:7a9071697d88ea7f25adfee43f579cb0012a79581c74d883e4ec8cb5a19cd896

Observation 0b8b784a-e963-4eea-88f9-63e0b67a47c8 · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

AbsenceBench: Language Models Can't Tell What's Missing RWKV: Reinventing RNNs for the Transformer Era

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.744118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.744118Z digest=sha256:60b7694717e721eec22237d1445a0d6762e6e2de4e69823effc2a3a996f9d4e8

Observation f0555fb9-7405-4359-b81d-a649a992860b · outbound

This paper cites Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation.

AbsenceBench: Language Models Can't Tell What's Missing Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.747278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.747278Z digest=sha256:07953d578fc595913dc29d0571bd73c4822ada3dfa485732fd517672607eec17

Observation 0fb76f57-1b76-41a0-8d9c-66565c113efe · outbound

This paper cites Qwen3 technical report, 2025 a.

AbsenceBench: Language Models Can't Tell What's Missing Qwen3 technical report, 2025 a

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.202173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.751186Z digest=sha256:08225a2f2caab1d3b0cc74e0296ee7086474c44b079d7a072c99295c6531f61b

Observation 624fdfa1-57c7-4b31-aa60-6f6d80e6ebb9 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, 2025 b.

AbsenceBench: Language Models Can't Tell What's Missing Qwq-32b: Embracing the power of reinforcement learning, 2025 b

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.192878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.754563Z digest=sha256:10ff849b31e9c42ab10f98aa95e79033cad6107ed2c17dd906bc7df423450418

Observation 5ef853bc-f219-4f16-baf3-d60978a1bcf4 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

AbsenceBench: Language Models Can't Tell What's Missing Code Llama: Open Foundation Models for Code

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.757980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.757980Z digest=sha256:e961d16070a9c2733ac2fd5be7d6bd79a76a6c66c00068af3b219a603de70f1a

Observation 6d8bb6cf-7172-4920-9c7b-ee88dc9a0926 · outbound

This paper cites Z ero SCROLLS : A zero-shot benchmark for long text understanding.

AbsenceBench: Language Models Can't Tell What's Missing Z ero SCROLLS : A zero-shot benchmark for long text understanding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.761186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.761186Z digest=sha256:c830579acb32759eea87195188e408ef0ea18124075a63d8c8eb744822edb074

Observation 5f5985b6-1d6a-4717-94f6-42bb50b8cfe6 · outbound

This paper cites RoFormer: Enhanced Transformer with Rotary Position Embedding.

AbsenceBench: Language Models Can't Tell What's Missing RoFormer: Enhanced Transformer with Rotary Position Embedding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.764087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.764087Z digest=sha256:74a95f4f26fd700762e44c31aa02fa7dd7488417c980a6a375b0052b1532fce1

Observation 779560ff-d125-4b76-a476-ab4d55c1d77c · outbound

This paper cites A Length-Extrapolatable Transformer.

AbsenceBench: Language Models Can't Tell What's Missing A Length-Extrapolatable Transformer

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.767285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.767285Z digest=sha256:aed0934dbe90a510aa0fc7dab0ead20956ab59eccc223c555293d1c773fdd47a

Observation ef0c8129-be67-40d4-8a09-820e953193ad · outbound

This paper cites Attention is all you need.

AbsenceBench: Language Models Can't Tell What's Missing Attention is all you need

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.770366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.770366Z digest=sha256:f4898ba67d2da127f7c8305936f99b7c892696cda828199ad789a2461716c543

Observation 41be1d9e-0019-4b67-bfaf-49151455a136 · outbound

This paper cites Michelangelo: Long Context Evaluations Beyond Haystacks via Latent Structure Queries.

AbsenceBench: Language Models Can't Tell What's Missing Michelangelo: Long Context Evaluations Beyond Haystacks via Latent Structure Queries

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.773223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.773223Z digest=sha256:959dc01145ac96824636513d089987b40b344be8f33106bf7ebf86d7f78e931a

Observation f6b28314-4984-4097-8f42-ffe392ba4c02 · outbound

This paper cites NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens.

AbsenceBench: Language Models Can't Tell What's Missing NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.776432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.776432Z digest=sha256:7b0c6411be67dc212a2ac94b31930aa6b928050afc1a5df8e727f295aa2c5c52

Observation e8c5b1f1-76a0-45ce-9e64-548cd4be8f3a · outbound

This paper cites Grok 3 beta — the age of reasoning agents, 2025.

AbsenceBench: Language Models Can't Tell What's Missing Grok 3 beta — the age of reasoning agents, 2025

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.177910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.779538Z digest=sha256:047dd70294e941c01392fe855e807bba5cbe65b6bc4e666cfbedc126763bdd56

Observation 77426410-528d-43a0-a526-55cec043c59f · outbound

This paper cites Stress-Testing Long-Context Language Models with Lifelong ICL and Task Haystack.

AbsenceBench: Language Models Can't Tell What's Missing Stress-Testing Long-Context Language Models with Lifelong ICL and Task Haystack

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.782387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.782387Z digest=sha256:d4c090021d67e42d8c4da8e6b107169ef242f5b5dcd80be16d611f8d03963397

Observation 0416608e-094e-4f4f-a4c2-4a20a02e5fc6 · outbound

This paper cites Qwen2.5 Technical Report.

AbsenceBench: Language Models Can't Tell What's Missing Qwen2.5 Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.785799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.785799Z digest=sha256:851eddb480171594f8bc3e7ee2a6d637961545ea30ea124cce6e1c13403485fd

Observation 102efd5b-655b-4ab6-ba39-f87b1ae4a122 · outbound

This paper cites Long-Context Language Modeling with Parallel Context Encoding.

AbsenceBench: Language Models Can't Tell What's Missing Long-Context Language Modeling with Parallel Context Encoding

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.788599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.788599Z digest=sha256:13a3fb2940de0131502fff2045465cfeb7da9ce2d0116fed09b0021186889561

Observation 23303ff9-36c3-4125-bd84-9493b8df1701 · outbound

This paper cites HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly.

AbsenceBench: Language Models Can't Tell What's Missing HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.791605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.791605Z digest=sha256:c4e8558b2f42736d7bd3eaf9a65627fbd2632d74f7d894f14f0a0cdcd95db90b

Observation f936a41a-b2a4-4d44-b710-f6820d013140 · outbound

This paper cites $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens.

AbsenceBench: Language Models Can't Tell What's Missing $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.794692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.794692Z digest=sha256:7fb64ebc60a6d9f072f9c21611a013e991b16e6dc44a8a98bef05c6edb42186d

Observation 0e55bcf3-5a6b-4018-9736-12d0dcdedbaf · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

AbsenceBench: Language Models Can't Tell What's Missing Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.797667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.797667Z digest=sha256:8850f420e33114eea208c97fccd0edeb3284f7aa8cf822a68fb2205884b9df49

Pith citing papers

Observation a529bcd7-92a9-4507-b7d0-d067eb575c35 · inbound

Mitigating hallucinations and omissions in LLMs for invertible problems: An application to hardware logic design automation cites this paper.

Mitigating hallucinations and omissions in LLMs for invertible problems: An application to hardware logic design automation AbsenceBench: Language Models Can't Tell What's Missing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:29:05.107544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-17T05:24:05.830241Z digest=sha256:fdc937dba6bdf62b969ce43e614502159c65868f94000531f4154d8af63d4c2f

Observation d5bceb12-34fd-4946-9055-885ddeb8c53e · inbound

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems cites this paper.

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems AbsenceBench: Language Models Can't Tell What's Missing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:31:17.313713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T15:12:00.081798Z digest=sha256:c831c6bc278f06dbc9123c6091c47659f7665edb78d3ebd238056b652b96f653

Observation 27e3d2c1-b4be-484d-9a5b-c41848ca5ebf · inbound

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems cites this paper.

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems AbsenceBench: Language Models Can't Tell What's Missing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:35:25.217000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-25T06:31:00.249428Z digest=sha256:d66c3a505dee930bc64a101564421c12241a856108874809edc578053a51c66b

Observation 85b9209b-8f41-477c-a373-1c1bcf0298e8 · inbound

Locality Does Not Imply Reachability: Boundary Repair in Block-Sparse Causal Attention cites this paper.

Locality Does Not Imply Reachability: Boundary Repair in Block-Sparse Causal Attention AbsenceBench: Language Models Can't Tell What's Missing

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:16:16.395398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T15:34:46.265644Z digest=sha256:ac51e2791d42daabb4f8c81262a1d414dbcc0c0670843881fcfe82c5c9df9675

Observation 3c3b00df-813b-4824-b32f-191aab9e1cd9 · inbound

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals cites this paper.

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:19:50.247302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T05:24:11.228672Z digest=sha256:889f03418ad00bfb9c0ff58e76b3c9811da819ca1e8c86b41d1397da78c823cf

Observation de7b2486-b4fc-40ae-9dac-3c900777e07b · inbound

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals cites this paper.

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T11:58:52.700648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:58:52.700648Z digest=sha256:eea1c1c11ebf1cd52cc6f33a603db4013ab8c600cd66f76ab19c3acdf5f71751

Observation c04a9899-bad2-4526-b63f-f5d0e9e2010e · inbound

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals cites this paper.

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T04:42:46.166372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:42:46.166372Z digest=sha256:585d44c73243fb59328b43f58b8cb164eb095b978823c0aad4e8b49db5a6a93a

Observation f0cc3f5a-3a52-4705-8a0d-733ca1a87e7f · inbound

When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models cites this paper.

When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models AbsenceBench: Language Models Can't Tell What's Missing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:22:44.774647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:22:44.774647Z digest=sha256:a0d8b2426867b2f342eb44b5deae64ad1423ad097bcc728120954f06a9e066e2