Pith. sign in

Paper Citation Record · LEDGER

AbsenceBench: Language Models Can't Tell What's Missing

As of 8 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 8 inbound Pith citation observations for arXiv:2506.11440.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11440 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:13:17.797667Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:22:44.774647Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:19:50.245560Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9bd0fe80-df29-4d2b-a589-d3c6b0e632c2 · outbound

This paper cites Many-Shot In-Context Learning.

AbsenceBench: Language Models Can't Tell What's Missing Many-Shot In-Context Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.648452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.648452Z digest=sha256:5f7957e8f4e663b9059fb4488f3fede7d734c1fcfe4606b4caf9f51dc4a9ef54

Observation dd26a989-0335-450d-9d89-aa1080f9a8d6 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

AbsenceBench: Language Models Can't Tell What's Missing Yi: Open Foundation Models by 01.AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.652986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.652986Z digest=sha256:544db9d9b826b85fd986cbab81b331a2dee87744064cf031769ee28198bf944e

Observation 56e6fd7a-77b1-4878-bb0d-11a3c679c1f8 · outbound

This paper cites L-Eval: Instituting Standardized Evaluation for Long Context Language Models.

AbsenceBench: Language Models Can't Tell What's Missing L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.656859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.656859Z digest=sha256:adfe361d088f8d3a469ce030abc4e2652722d39b54961e4564ac3dcf58c5c465

Observation cdcc8374-c029-4222-ba5e-64aaaac30c8c · outbound

This paper cites Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud.

AbsenceBench: Language Models Can't Tell What's Missing Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.265746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.660267Z digest=sha256:dbc7b888fafec0ee05373ef6fdba0394a0f81840e3a3df8ee86d16693a47bebf

Observation 59e6c12b-1be1-44d5-aa28-03d6df593380 · outbound

This paper cites Claude 3 haiku: our fastest model yet, 2024.

AbsenceBench: Language Models Can't Tell What's Missing Claude 3 haiku: our fastest model yet, 2024

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.256396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.663450Z digest=sha256:886f095f67e4a1638a7605ada3acf94c0d2f448a403fe95559740dfa3e714432

Observation 4b240ed1-6152-4706-af7c-a8f15e1ee392 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

AbsenceBench: Language Models Can't Tell What's Missing LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.666475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.666475Z digest=sha256:98d4a1223457a8095d6705dbd20a9f625640b7052400582b6beab165250437af

Observation 26baed68-fc15-4741-9621-68208e1e5924 · outbound

This paper cites Unlimiformer: Long-Range Transformers with Unlimited Length Input.

AbsenceBench: Language Models Can't Tell What's Missing Unlimiformer: Long-Range Transformers with Unlimited Length Input

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.670308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.670308Z digest=sha256:3b4f55c56d4758f721b866c3065989ac7f5eedb4760493420b361cd517b6e1df

Observation b11f4f4b-22c4-454e-8ca2-8d6c6f6f1730 · outbound

This paper cites BooookScore: A systematic exploration of book-length summarization in the era of LLMs.

AbsenceBench: Language Models Can't Tell What's Missing BooookScore: A systematic exploration of book-length summarization in the era of LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.673581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.673581Z digest=sha256:fa269c86374beb56c9036490495414fb5c0d6b564006028f448d3762851df28b

Observation 6acd2405-5ee4-40df-ae14-56d6fbbad1a0 · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

AbsenceBench: Language Models Can't Tell What's Missing Extending Context Window of Large Language Models via Positional Interpolation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.676871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.676871Z digest=sha256:90d461c64f6c1e79e533909f61a5b00d2ad57e0a70c5f1f4441f4215df1e32a6

Observation b4aeb2d8-1861-4e57-b56d-81de0cdde438 · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

AbsenceBench: Language Models Can't Tell What's Missing Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.679986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.679986Z digest=sha256:92550c7c0f57a27c2db4b8afd67fb212fe6d1560bada65dbb0bf0f57de46cbd6

Observation 86d5eac7-5bc7-4027-b66b-cf8d190b1f3c · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AbsenceBench: Language Models Can't Tell What's Missing DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.683266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.683266Z digest=sha256:d3552840567748007ff10e6ebd61725f63e10f346b755d54ef2facd560255c06

Observation ce77bd0a-7837-4a4f-9fcd-ff3e30cd340a · outbound

This paper cites Mathematical Capabilities of ChatGPT.

AbsenceBench: Language Models Can't Tell What's Missing Mathematical Capabilities of ChatGPT

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.686818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.686818Z digest=sha256:f613abe1edb730cbf520be40617789fc8b2e2b0ef1d63128a1e073a7ec1c25a7

Observation 8c8f867e-c763-44a5-b07b-fafc618db48c · outbound

This paper cites Simple Hardware-Efficient Long Convolutions for Sequence Modeling.

AbsenceBench: Language Models Can't Tell What's Missing Simple Hardware-Efficient Long Convolutions for Sequence Modeling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.690031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.690031Z digest=sha256:64247d5787003637ab04ed0b384b6b8fa25fcac708143146d87fcaeab2df9d6f

Observation 91348c31-eb1e-4342-aa92-49a2c48aff3f · outbound

This paper cites How to train long-context language models (effectively), 2025.

AbsenceBench: Language Models Can't Tell What's Missing How to train long-context language models (effectively), 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.693121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.693121Z digest=sha256:de406e09ae4a0554f14449bc9d675e237b1da43d61670a0b563c17798103ee3b

Observation 314cd94a-af43-44cc-b755-1474571bdac2 · outbound

This paper cites Gemini 2.5: Our most intelligent ai model, 2025.

AbsenceBench: Language Models Can't Tell What's Missing Gemini 2.5: Our most intelligent ai model, 2025

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.247436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.695912Z digest=sha256:242fd7763f4e5205ab302a3a2d0724f32fe8b529130d2d88b1ccc2311ca47b78

Observation 3a79523c-c853-49ec-a913-ac6ebc8bd7ff · outbound

This paper cites The Llama 3 Herd of Models.

AbsenceBench: Language Models Can't Tell What's Missing The Llama 3 Herd of Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.699059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.699059Z digest=sha256:0504c9ab56a28611d2179c5727f83a415b447e69874067e5fcbe7d13cdf9a96a

Observation 95615ecf-e4c3-4f3f-afe8-9f3d94364ce2 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

AbsenceBench: Language Models Can't Tell What's Missing Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.702868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.702868Z digest=sha256:52f920f0f3e296ad438c5adaac3f723c5454c779180f9e0eb807cafb97514830

Observation 534072d9-ab58-442f-a018-96860604436e · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

AbsenceBench: Language Models Can't Tell What's Missing RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.705952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.705952Z digest=sha256:8a5007a56ad68de9696348487d00d3c2a5a688e7f2b08ce8f4df1446434db3b5

Observation 23c9b130-cb9c-4e0c-97d7-b9063af6f734 · outbound

This paper cites Mixtral of Experts.

AbsenceBench: Language Models Can't Tell What's Missing Mixtral of Experts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.709044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.709044Z digest=sha256:33fb99c52e49fa667085f7e86cf83e7abd112d180160c3434bccc5f137d71128

Observation 984be04b-3503-41a5-b086-912982775004 · outbound

This paper cites Needle in a haystack - pressure testing llms, 2023.

AbsenceBench: Language Models Can't Tell What's Missing Needle in a haystack - pressure testing llms, 2023

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.238761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.712296Z digest=sha256:b2f770829b4b02b6f413bae75d31c2f1f1db63cda4e5fd6fbcfb5b7ed9378179

Observation 23c904a3-feed-4302-b6b7-fc1b8b932fd3 · outbound

This paper cites FABLES: Evaluating faithfulness and content selection in book-length summarization.

AbsenceBench: Language Models Can't Tell What's Missing FABLES: Evaluating faithfulness and content selection in book-length summarization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.715567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.715567Z digest=sha256:a7deb13bc316dd5d2eb1b0c432dc7624b1a349e11083550068353f73e73a9fa9

Observation 6c134385-792b-4c51-9ea6-155b1800f4e3 · outbound

This paper cites Benchmarking Cognitive Biases in Large Language Models as Evaluators.

AbsenceBench: Language Models Can't Tell What's Missing Benchmarking Cognitive Biases in Large Language Models as Evaluators

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.718512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.718512Z digest=sha256:c6db4c5657991d6a9f9fb2c66cd4b677d5ec9c5f5ec63b7232a89a22bddd339a

Observation 3e5344b0-c219-46e2-99c1-86b488710e7e · outbound

This paper cites The NarrativeQA Reading Comprehension Challenge.

AbsenceBench: Language Models Can't Tell What's Missing The NarrativeQA Reading Comprehension Challenge

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.721714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.721714Z digest=sha256:5afc9aaf5c4d73a08811a41b96dafde597a2b69b7b25cdb8e6e0308adebf9a35

Observation 7bffc137-3dd6-44fc-9766-b3b995853ebb · outbound

This paper cites Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?.

AbsenceBench: Language Models Can't Tell What's Missing Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.725291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.725291Z digest=sha256:9ef088cfab13e5ddb0fce212c45ef38f05cbcf1a9466468b0d4dba40ed17ff61

Observation 0f02fd82-05aa-4263-8a70-41a75902de14 · outbound

This paper cites The llama 4 herd: The beginning of a new era of natively multimodal ai innovation, 2025.

AbsenceBench: Language Models Can't Tell What's Missing The llama 4 herd: The beginning of a new era of natively multimodal ai innovation, 2025

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.228912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.728350Z digest=sha256:ac990b73610f469e646a91144d94f897849ac4ab38132f36832080512bfba36a

Observation 48a5b14c-7bf1-4424-8f5f-cd326fc48a71 · outbound

This paper cites Openai o3-mini, pushing the frontier of cost-effective reasoning., 2025.

AbsenceBench: Language Models Can't Tell What's Missing Openai o3-mini, pushing the frontier of cost-effective reasoning., 2025

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.219852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.731389Z digest=sha256:872442644abfd36651c7d001b2b094756f462bd953df59e5b9078d7cd21c1954

Observation 1f0529d3-daf8-4b7f-8057-ad04af79bf06 · outbound

This paper cites GPT-4o System Card.

AbsenceBench: Language Models Can't Tell What's Missing GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.734425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.734425Z digest=sha256:9e4e9d927d6ac5757564c3b8a523a58d3a630d92aabc769a52f03dedf1ef049c

Observation 9e74647e-5850-4944-af1d-86d0873be01a · outbound

This paper cites GPT-4 Technical Report.

AbsenceBench: Language Models Can't Tell What's Missing GPT-4 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.737879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.737879Z digest=sha256:0f37e792ba2103c770eee8210a4f68a8395de1f757a0ee8668fc97b9bcd5aaac

Observation 9ccfa6f8-8ac4-4fc3-a740-588bac6e9edd · outbound

This paper cites gutenberg-poetry-corpus: A corpus of poetry from project gutenberg, 2018.

AbsenceBench: Language Models Can't Tell What's Missing gutenberg-poetry-corpus: A corpus of poetry from project gutenberg, 2018

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.210950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.741220Z digest=sha256:ac480cb6babe2c2e59011b6dce216c288605b695ac21ff1e0f54486ddea8d9b2

Observation 0b8b784a-e963-4eea-88f9-63e0b67a47c8 · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

AbsenceBench: Language Models Can't Tell What's Missing RWKV: Reinventing RNNs for the Transformer Era

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.744118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.744118Z digest=sha256:f4d6e9db013535215c646908513f85ec7a5942d03683cf3ef1125d87982876e5

Observation f0555fb9-7405-4359-b81d-a649a992860b · outbound

This paper cites Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation.

AbsenceBench: Language Models Can't Tell What's Missing Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.747278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.747278Z digest=sha256:7df7997e502e8c1f7b74b0447dd77a521ab036dd6534a7f1b240a8c986d967cc

Observation 0fb76f57-1b76-41a0-8d9c-66565c113efe · outbound

This paper cites Qwen3 technical report, 2025 a.

AbsenceBench: Language Models Can't Tell What's Missing Qwen3 technical report, 2025 a

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.202173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.751186Z digest=sha256:f8da1869847193b122efa33dcd9dab9335898aefe7fbfe8c7e44fb161b6ce81e

Observation 624fdfa1-57c7-4b31-aa60-6f6d80e6ebb9 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, 2025 b.

AbsenceBench: Language Models Can't Tell What's Missing Qwq-32b: Embracing the power of reinforcement learning, 2025 b

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.192878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.754563Z digest=sha256:2c721be78f264e922d6e0e10e6c08549e42541fb10ebe7dd728fd5fcca238378

Observation 5ef853bc-f219-4f16-baf3-d60978a1bcf4 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

AbsenceBench: Language Models Can't Tell What's Missing Code Llama: Open Foundation Models for Code

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.757980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.757980Z digest=sha256:2942457facb06a9ed7624dd79de3cb498c3c4bf2e28a2f76bb23a97a06c988e0

Observation 6d8bb6cf-7172-4920-9c7b-ee88dc9a0926 · outbound

This paper cites Z ero SCROLLS : A zero-shot benchmark for long text understanding.

AbsenceBench: Language Models Can't Tell What's Missing Z ero SCROLLS : A zero-shot benchmark for long text understanding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.761186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.761186Z digest=sha256:31506258086ea6c928ce2bbc7eb61ea6833ff350b80e1b215897ebc3426f6826

Observation 5f5985b6-1d6a-4717-94f6-42bb50b8cfe6 · outbound

This paper cites RoFormer: Enhanced Transformer with Rotary Position Embedding.

AbsenceBench: Language Models Can't Tell What's Missing RoFormer: Enhanced Transformer with Rotary Position Embedding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.764087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.764087Z digest=sha256:7850805445f6d2f233d8f6733d9de51fef55d4be3eb78d30cc378c08d53d3403

Observation 779560ff-d125-4b76-a476-ab4d55c1d77c · outbound

This paper cites A Length-Extrapolatable Transformer.

AbsenceBench: Language Models Can't Tell What's Missing A Length-Extrapolatable Transformer

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.767285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.767285Z digest=sha256:a5b84ba61ecf51d97904df8ffd79e9d6efed569aa2107c6b74962cbfdab302ba

Observation ef0c8129-be67-40d4-8a09-820e953193ad · outbound

This paper cites Attention is all you need.

AbsenceBench: Language Models Can't Tell What's Missing Attention is all you need

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.770366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.770366Z digest=sha256:aeb8ca0a5a51034d81b8aafaa7bacc8733d014217b73f4f9cc6c27171c46ce18

Observation 41be1d9e-0019-4b67-bfaf-49151455a136 · outbound

This paper cites Michelangelo: Long Context Evaluations Beyond Haystacks via Latent Structure Queries.

AbsenceBench: Language Models Can't Tell What's Missing Michelangelo: Long Context Evaluations Beyond Haystacks via Latent Structure Queries

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.773223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.773223Z digest=sha256:fde15c1dddd297f5b9ae14b53ec6e1102b522d8e70af345ea1c4ab0b0b5c23de

Observation f6b28314-4984-4097-8f42-ffe392ba4c02 · outbound

This paper cites NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens.

AbsenceBench: Language Models Can't Tell What's Missing NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.776432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.776432Z digest=sha256:f6436a6410b788b762e76b6c89c47d7650e9f2c4668eaac8ed34a55fec2fb38a

Observation e8c5b1f1-76a0-45ce-9e64-548cd4be8f3a · outbound

This paper cites Grok 3 beta — the age of reasoning agents, 2025.

AbsenceBench: Language Models Can't Tell What's Missing Grok 3 beta — the age of reasoning agents, 2025

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:13:18.177910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:13:17.779538Z digest=sha256:a55daf2fc8fd85580538b3d363c748de6e8cf07c9b40a630e019ac88545a5175

Observation 77426410-528d-43a0-a526-55cec043c59f · outbound

This paper cites Stress-Testing Long-Context Language Models with Lifelong ICL and Task Haystack.

AbsenceBench: Language Models Can't Tell What's Missing Stress-Testing Long-Context Language Models with Lifelong ICL and Task Haystack

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.782387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.782387Z digest=sha256:513910747af69806b4d2a7e170994562dc681d836e3cc41f682e6a9cc7d54fde

Observation 0416608e-094e-4f4f-a4c2-4a20a02e5fc6 · outbound

This paper cites Qwen2.5 Technical Report.

AbsenceBench: Language Models Can't Tell What's Missing Qwen2.5 Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.785799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.785799Z digest=sha256:04d06b8eb2a8c5bfa96223f77029f31ecaadbabcb91c1c4bbe72fed8c6915d99

Observation 102efd5b-655b-4ab6-ba39-f87b1ae4a122 · outbound

This paper cites Long-Context Language Modeling with Parallel Context Encoding.

AbsenceBench: Language Models Can't Tell What's Missing Long-Context Language Modeling with Parallel Context Encoding

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.788599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.788599Z digest=sha256:1affc356e2c5182330fb738cb168c9ebf6390d2354bf2b606cff206b97566ffe

Observation 23303ff9-36c3-4125-bd84-9493b8df1701 · outbound

This paper cites HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly.

AbsenceBench: Language Models Can't Tell What's Missing HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.791605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.791605Z digest=sha256:ed99c2c812cde810512623cedbe3640f9fcaf9df66beadc6a2939293ddce126e

Observation f936a41a-b2a4-4d44-b710-f6820d013140 · outbound

This paper cites $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens.

AbsenceBench: Language Models Can't Tell What's Missing $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.794692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.794692Z digest=sha256:95b25a16a35a5830ea3c3deb8d9c17518e6aa4f5a0381fbc2e8fc8fe5c213b05

Observation 0e55bcf3-5a6b-4018-9736-12d0dcdedbaf · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

AbsenceBench: Language Models Can't Tell What's Missing Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.797667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.797667Z digest=sha256:48bcffad2d698944c9f3415f37ff528292f57d0febe5206098f55b6d799dd3a8

Pith citing papers

Observation a529bcd7-92a9-4507-b7d0-d067eb575c35 · inbound

Mitigating hallucinations and omissions in LLMs for invertible problems: An application to hardware logic design automation cites this paper.

Mitigating hallucinations and omissions in LLMs for invertible problems: An application to hardware logic design automation AbsenceBench: Language Models Can't Tell What's Missing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:29:05.107544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T05:24:05.830241Z digest=sha256:2e8ed39b8a70e8a459153193a8f28d5bdeb2faf23c37dac5499569360eb20add

Observation d5bceb12-34fd-4946-9055-885ddeb8c53e · inbound

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems cites this paper.

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems AbsenceBench: Language Models Can't Tell What's Missing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:31:17.313713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T15:12:00.081798Z digest=sha256:d8ab15526b35cf7668a6ad5bc1b5004a002ac8b71120e4a7df68822a890589cf

Observation 27e3d2c1-b4be-484d-9a5b-c41848ca5ebf · inbound

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems cites this paper.

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems AbsenceBench: Language Models Can't Tell What's Missing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:35:25.217000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T06:31:00.249428Z digest=sha256:f5cfff70971e656b35a6e2dc05e4394d2bb97590f767e8a57743a00af571aeb6

Observation 85b9209b-8f41-477c-a373-1c1bcf0298e8 · inbound

Locality Does Not Imply Reachability: Boundary Repair in Block-Sparse Causal Attention cites this paper.

Locality Does Not Imply Reachability: Boundary Repair in Block-Sparse Causal Attention AbsenceBench: Language Models Can't Tell What's Missing

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:16:16.395398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T15:34:46.265644Z digest=sha256:53515baa538a95fdfdd587967be5706e14d3b131769e8b38283a7d2303e29171

Observation 3c3b00df-813b-4824-b32f-191aab9e1cd9 · inbound

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals cites this paper.

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:19:50.247302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:24:11.228672Z digest=sha256:50f50515dbb898604cdd24c193736287037ea64b472cba4ae3610bbb9f992f51

Observation de7b2486-b4fc-40ae-9dac-3c900777e07b · inbound

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals cites this paper.

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T11:58:52.700648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:58:52.700648Z digest=sha256:e85fc8cc4c3eaa831331b0de4b8826a333f6c76606f6d35ea1e8fbbe7ed9fba7

Observation c04a9899-bad2-4526-b63f-f5d0e9e2010e · inbound

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals cites this paper.

The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals AbsenceBench: Language Models Can't Tell What's Missing

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T04:42:46.166372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:42:46.166372Z digest=sha256:a01605114fecc924348cd7a5ad3d78161a2053e0741ed1d726a9133845b5b210

Observation f0cc3f5a-3a52-4705-8a0d-733ca1a87e7f · inbound

When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models cites this paper.

When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models AbsenceBench: Language Models Can't Tell What's Missing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:22:44.774647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:22:44.774647Z digest=sha256:da4c27d922f7c6ce591fa256fa8cd1c8644d4ebc322887b02106370b8262ed9d