Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:46:27.690613Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2507.22411.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:46:27.690613Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 06c60503-8b14-4bb6-8ee8-de3c0d9b0744 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models L -eval: Instituting standardized evaluation for long context language models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6705ba03-5846-494e-b84d-59c35f98ae6b · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Why Does the Effective Context Length of LLMs Fall Short?
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d49ffe9-54d1-4eb0-8f97-585cb830a97e · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models L ong B ench: A bilingual, multitask benchmark for long context understanding
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 828392cb-c7af-4d80-94da-ee230c7bd4c0 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Lost in the haystack: Smaller needles are more difficult for llms to find
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9b5b160-22e3-4aa3-ac54-f185959fde41 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Longrope: extending llm context window beyond 2 million tokens
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7c47c35-0e6b-4924-97db-9975f380f9ff · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Robinson, Keren Gu, Anna-Luisa Brakman, Pamela Mishkin, Meghan Shah, Johannes Heidecke, Lilian Weng, and Adam Tauman Kalai
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 85064a81-8709-48d6-8bbf-ed26bf4b26d8 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models A little goes a long way: Efficient long context training and inference with partial contexts
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dbe909cd-9916-41f0-a0a4-99660e694531 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models The Llama 3 Herd of Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fafd8105-793c-48ff-a563-0cc581de2123 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d01add48-d443-4b30-8a68-8236896c9903 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models GPT-4o System Card
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bd5ee6b-52da-4fa4-ad8f-8ca1d8b9db28 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Llm maybe longlm: Selfextend llm context window without tuning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 036587d2-ef5d-4365-be44-8c051fdd258c · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Babilong: Testing the limits of llms with long context reasoning-in-a-haystack
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3d7013a5-3ff9-46b1-bf89-dd6fe1a282e5 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Efficient memory management for large language model serving with pagedattention
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3716063-93bb-4964-a404-3f4fec1d63ba · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Summary of a haystack: A challenge to long-context llms and rag systems
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cfca3c99-9896-42a2-a3ef-ea2e42b3d7bd · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63b56fc7-333b-4e53-ac73-935f888ff4aa · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Needlebench: Can llms do retrieval and reasoning in 1 million context window? arXiv preprint arXiv:2407.11963, 2024 b
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 522563eb-6d57-4c54-a6be-7c4bcbd31790 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models MARIO : MA th reasoning with code interpreter output - a reproducible pipeline
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf94dfeb-a0f7-4c58-a42f-4f747c5f917d · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models DeepSeek-V3 Technical Report
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05a0efc1-82cd-4def-9a61-b1d2f81b6207 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Lost in the middle: How language models use long contexts
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 34b04707-0861-4270-805c-30383b7df666 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models The llama 4 herd: The beginning of a new era of natively multimodal ai innovation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 654112ab-d448-4106-9f0f-60c86e5bb4f2 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Ya RN : Efficient context window extension of large language models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f15e1a9e-cbb7-4de4-8082-a7cba26892b7 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models o ring, and Julius Tr \
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4d393cd8-b644-4cf8-8f21-7d5f9c4983bd · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models G eo C oder: Solving geometry problems by generating modular code through vision-language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4c602030-6238-457c-a450-0f5b6b115e8f · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Counting-stars: A multi-evidence, position-aware, and scalable benchmark for evaluating long-context large language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e333c5d0-ee7e-4b50-9ac6-b73ee2f45e6b · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Gemma 3 Technical Report
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c83896ea-3553-4513-a4a9-d68fc873aa95 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Qwq-32b: Embracing the power of reinforcement learning, 2025
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffe97308-7504-44dc-925e-23f1cfda8f71 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9144b23-141c-406c-bca0-736ba4a2905c · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Focused transformer: Contrastive training for context scaling
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7bb3c8db-7d2c-4260-b2ef-d648aaa7a923 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Needle in a multimodal haystack
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4e8a3c6f-08f0-4d1a-85e7-4503d2cc7f49 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Transformers: State-of-the-art natural language processing
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2ccb463-1a52-473d-86e5-75f84a56734b · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Qwen2.5 Technical Report
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af51c293-28e1-4759-b51a-8e37ff4afe55 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Qwen3 Technical Report
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74fd94d5-9007-46ef-a4ca-df8740aed517 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Qwen2.5-1M Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8559064b-2315-4680-8853-a03504f67425 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Sequential-niah: A needle-in-a-haystack benchmark for extracting sequential needles from long contexts
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e288130a-9477-44aa-8fe4-4c0d804077bc · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models B ench: Extending long context evaluation beyond 100 K tokens
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2984a17-bc41-403c-857b-58a98bdab347 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models write newline
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b353e21d-8a58-4bd4-be83-64542fdf0e0b · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models @esa (Ref
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90a33ccb-bd36-4710-82e5-3efc165e196e · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2661babe-2179-4fb7-b73e-f640903c2a28 · outbound
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models ROPE Contraction
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.