Pith. sign in

Paper Citation Record · LEDGER

NoLiMa: Long-Context Evaluation Beyond Literal Matching

As of 9 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 28 inbound Pith citation observations for arXiv:2502.05167.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.05167 v3

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T20:07:33.797188Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:22:17.501214Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

4
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 2a14e8e5-e940-4aa9-8355-6d9095d382ee · outbound

This paper cites write newline.

NoLiMa: Long-Context Evaluation Beyond Literal Matching write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.592598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.592598Z digest=sha256:9352f108bdaf60d6bb17d07bafcfb2310b0b3c7eeb83b3e5d21f8f0f5ce5060e

Observation 173c4c31-33e3-41c3-8177-b63e5216be83 · outbound

This paper cites M., Bohnet, B., Rosias, L., Chan, S.

NoLiMa: Long-Context Evaluation Beyond Literal Matching M., Bohnet, B., Rosias, L., Chan, S

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.642362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.598942Z digest=sha256:8da65700527c97a603e8eba56a7b424c5d1ef4cbe0eaac94a328a17fb7bf7336

Observation f5299136-3f17-410b-be2d-6027c9914ce7 · outbound

This paper cites Claude 3.5 sonnet model card addendum.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Claude 3.5 sonnet model card addendum

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.603825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.603825Z digest=sha256:d5ade40a2f68dad36ca94928ca1635c1827659716052262b185fa36aca86a349

Observation a998d468-284c-4cef-9727-eec3fa264fdb · outbound

This paper cites Zoology: Measuring and improving recall in efficient language models.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Zoology: Measuring and improving recall in efficient language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.616107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.608952Z digest=sha256:3e104ec58f7d1c010a932a55496f500f59ef56e9c10c7aae877ea8d25d6c765e

Observation 54cf8e90-7a22-4c59-9698-9901626cefed · outbound

This paper cites E., Mnih, V., Leibo, J.

NoLiMa: Long-Context Evaluation Beyond Literal Matching E., Mnih, V., Leibo, J

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.613359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.613359Z digest=sha256:2ef9d1cc6f7731bb9d057298b51032d8c07c3d10afe0ca2b7ac520fc0571e677

Observation 9dd5cff6-1f4b-4b5e-add3-1eca2b3e40b8 · outbound

This paper cites L ong B ench: A bilingual, multitask benchmark for long context understanding.

NoLiMa: Long-Context Evaluation Beyond Literal Matching L ong B ench: A bilingual, multitask benchmark for long context understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.617862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.617862Z digest=sha256:77cd0744bba2b3a750ca2741672fee78ce0d168e44c49554dcec926f1f710710

Observation fc3f35dd-9072-4cd8-ac73-51d5643ce929 · outbound

This paper cites Booookscore: A systematic exploration of book-length summarization in the era of LLM s.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Booookscore: A systematic exploration of book-length summarization in the era of LLM s

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.589134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.622621Z digest=sha256:608a227e1df91ea16bdafd08d3882030e61cf7e9cc2d971933ea005c6c720566

Observation 29bb0b64-13fb-4439-a9d7-614c8e44e91b · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Extending Context Window of Large Language Models via Positional Interpolation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.627669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.627669Z digest=sha256:e6d43f75d79c470f35b2f3d24e160d2da1486f249f738cfed8585410d5286419

Observation 6735a852-b89f-4106-b095-89e10c83012b · outbound

This paper cites c4ai-command-r-plus-08-2024, 2024.

NoLiMa: Long-Context Evaluation Beyond Literal Matching c4ai-command-r-plus-08-2024, 2024

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.573388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.632317Z digest=sha256:47ff6fe6814e78caaafbbbefe88e26357881f5ec672a3060d5cc9489be9214cb

Observation fdaa6f2e-312c-4c9b-8bb2-468681b095c5 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

NoLiMa: Long-Context Evaluation Beyond Literal Matching DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.636486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.636486Z digest=sha256:9c78af0c32cea8a227c27c258fd7a35279a66f734c0b93bec1d92a362623bd34

Observation 124b99d0-b237-4587-b0f3-a6e86130ca78 · outbound

This paper cites X., and Wen, J.-R.

NoLiMa: Long-Context Evaluation Beyond Literal Matching X., and Wen, J.-R

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.557958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.641111Z digest=sha256:3fa47e01293fc3fdd2e4d640807bc4eab3e76b41b510c970ad62afd4fee2cf64

Observation e3fb69c7-51d9-4256-922e-688b34946c76 · outbound

This paper cites The Llama 3 Herd of Models.

NoLiMa: Long-Context Evaluation Beyond Literal Matching The Llama 3 Herd of Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.645488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.645488Z digest=sha256:1b819b6a776b20949b09be204bab0e6ce02543199e3281f3b11fde7442d75bb7

Observation e167a8a6-8813-4263-8588-15969d5b02b7 · outbound

This paper cites Is it really long context if all you need is retrieval? towards genuinely difficult long context NLP.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Is it really long context if all you need is retrieval? towards genuinely difficult long context NLP

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.649807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.649807Z digest=sha256:4467a72b12e4c3842df95480de0f574a8e8dc4396626d4608a9b2374cee7b931

Observation e593d0f3-309c-478e-bcc3-cef08e0a1cf3 · outbound

This paper cites Neural Turing Machines.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Neural Turing Machines

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.654022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.654022Z digest=sha256:bbd84068f723a2eee18004f3505a5894718170d5e426a6e1e87c083ea40185ab

Observation 1c05ec9c-879e-4756-9d04-ba84c2b8984c · outbound

This paper cites Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.658485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.658485Z digest=sha256:5c4b8b114e56a4bfe16f4561fa78f7e3276b59d073804b95007deb24a92bd477

Observation 3e3c721b-caad-46b8-a5c5-a33f7efa2d4a · outbound

This paper cites RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024.

NoLiMa: Long-Context Evaluation Beyond Literal Matching RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.663325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.663325Z digest=sha256:04e138c2bc6b8508ad3a151fff5804973f03970ce8d73027ab08d72fd57d5b7c

Observation 5f1399f4-2895-4cdd-a258-9ccaa8e5c1bd · outbound

This paper cites GPT-4o System Card.

NoLiMa: Long-Context Evaluation Beyond Literal Matching GPT-4o System Card

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.667798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.667798Z digest=sha256:725df7092847a603556a006e5483fcf86d35f1693ab4b02a238b1e932252a89e

Observation 54347c29-18a5-4c8c-84c2-9e1e822b28da · outbound

This paper cites Unsupervised dense information retrieval with contrastive learning.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Unsupervised dense information retrieval with contrastive learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.532417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.672794Z digest=sha256:7b430e293bb5a3aba406895e38d2176c14235075a16ddf07950fe1f24401365a

Observation a671bd00-4858-4793-816d-fc1b0e9bd21c · outbound

This paper cites J., Taylor, C.

NoLiMa: Long-Context Evaluation Beyond Literal Matching J., Taylor, C

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.677277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.677277Z digest=sha256:61d461b129165d04031b4f6150a61323840a8d82871be184eb6aa784d848f77e

Observation 2109e06f-d3ba-43db-abd2-4e64838141f4 · outbound

This paper cites Needle in a haystack-pressure testing llms.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Needle in a haystack-pressure testing llms

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.516136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.682949Z digest=sha256:65f284128858451f08e4ca64cef862280fc705c933a54028b59b6956225f7bd7

Observation 2765d5ce-25fa-4914-b20c-f8f223401e45 · outbound

This paper cites S., Reid, M., Matsuo, Y., and Iwasawa, Y.

NoLiMa: Long-Context Evaluation Beyond Literal Matching S., Reid, M., Matsuo, Y., and Iwasawa, Y

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.499918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.687443Z digest=sha256:d3b286205ce3c9a61ad4c7c283081bd450c6d3dc90076498c6e9939c20b28472

Observation 33b54eb3-ff1b-4d06-a727-27a311300fa0 · outbound

This paper cites I., Sorokin, A., and Burtsev, M.

NoLiMa: Long-Context Evaluation Beyond Literal Matching I., Sorokin, A., and Burtsev, M

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.484943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.691753Z digest=sha256:1091e9a4164977c168a48cab7e6d2ba14a28274ab546f2e056a4abc384105088

Observation a1d88382-70ec-43b2-b28d-7542a6d8e741 · outbound

This paper cites H., Gonzalez, J.

NoLiMa: Long-Context Evaluation Beyond Literal Matching H., Gonzalez, J

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.696215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.696215Z digest=sha256:31294a268e3d82436b79a3866d6fe502db20e34edd753f9cf83377a2d7af9657

Observation 74a1b7f6-5598-4845-bfb2-62b2bba3eced · outbound

This paper cites Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.700371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.700371Z digest=sha256:8a953622625052787edd3280427c0f32c0ce6b7782c4a0f9928dbc5b920df381

Observation 0b8bdee4-a411-45b7-ae77-26d0e427c0ee · outbound

This paper cites Same task, more tokens: the impact of input length on the reasoning performance of large language models.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Same task, more tokens: the impact of input length on the reasoning performance of large language models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.704728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.704728Z digest=sha256:3fe45601ecac1898aeb797bbd4e361187adbd9b16603d283eb95e67977ecc87a

Observation c6a152d2-ce70-478b-bffc-493308f22204 · outbound

This paper cites Needlebench: Can llms do retrieval and reasoning in 1 million context window?, 2024.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Needlebench: Can llms do retrieval and reasoning in 1 million context window?, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.709382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.709382Z digest=sha256:7031a7de8838b8053e19cf315342fe75558ad8239c338a71dd07abc4f25229b0

Observation 415d5cc6-337a-461e-8aac-dd758d1bba8b · outbound

This paper cites ROUGE : A package for automatic evaluation of summaries.

NoLiMa: Long-Context Evaluation Beyond Literal Matching ROUGE : A package for automatic evaluation of summaries

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.713886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.713886Z digest=sha256:d473357190c56407463ac8431ced99caa9bb8d0952fd662d05fe84448f9d157d

Observation 03602941-5f36-4afb-bdec-06d167f77f77 · outbound

This paper cites F., Lin, K., Hewitt, J., Paranjape, A., Bevilacqua, M., Petroni, F., and Liang, P.

NoLiMa: Long-Context Evaluation Beyond Literal Matching F., Lin, K., Hewitt, J., Paranjape, A., Bevilacqua, M., Petroni, F., and Liang, P

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.718306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.718306Z digest=sha256:93924686889e11183f40d43ccf1311e0f036a312a451be00b42a55a669976818

Observation 6c9ca1ef-a75a-4bb0-83d4-c6daba0b390d · outbound

This paper cites Evaluating very long-term conversational memory of LLM agents.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Evaluating very long-term conversational memory of LLM agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.722803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.722803Z digest=sha256:d0795d31629d6ece799ce83e32003a741b2387df52f8756d526ada970a8312ce

Observation 8d3d75cd-eccb-4f16-83b0-70004517f339 · outbound

This paper cites Llama 3.3 model card.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Llama 3.3 model card

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.448376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.727378Z digest=sha256:939a9d6b2d3a15a8702a056a9a78a690594097f46bc715b99b0a5a2dd42e5daa

Observation 6c9e4c0e-e819-49e5-afc9-7dce550a622e · outbound

This paper cites Mistral large 2.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Mistral large 2

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.433216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.731713Z digest=sha256:085efe75d2feaa6c8e52a6a6c5c04c55fc1c07032596a2f065b0c381b31ad864

Observation 858afd15-72f3-43e8-8eca-e23909d8a962 · outbound

This paper cites and Jaggi, M.

NoLiMa: Long-Context Evaluation Beyond Literal Matching and Jaggi, M

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.417927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.735841Z digest=sha256:eb73db71d4bd78dd1bd9312b87ef11e3476d7a2ae73a69ac0e85a45617097d82

Observation 54c0df92-392e-49ab-818d-a7af43f37745 · outbound

This paper cites Biases in large language models: Origins, inventory, and discussion.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Biases in large language models: Origins, inventory, and discussion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.739849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.739849Z digest=sha256:c0b61f15bd1fc62053e3fb4028f518f377df373fb4569d5caec0769703b86ba8

Observation ab5d9051-ec2e-425f-b55a-9fade2186663 · outbound

This paper cites In-context Learning and Induction Heads.

NoLiMa: Long-Context Evaluation Beyond Literal Matching In-context Learning and Induction Heads

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.744359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.744359Z digest=sha256:2b7f7b1f0e2f189fa71f9d95050e737d24d42aaa5a32f6db6a4390f215b0cc27

Observation e010c381-3a1c-46b0-8ed4-3132f81b9506 · outbound

This paper cites Openai o3-mini system card.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Openai o3-mini system card

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.403138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.748962Z digest=sha256:d1c9456d09aa98670ea8d33fb42fcbedaeb9e1edbad6e1af72dcb5d4df51e5ac

Observation 63e94e1e-3d22-4f09-9853-d24b343a0be0 · outbound

This paper cites OpenAI o1 System Card.

NoLiMa: Long-Context Evaluation Beyond Literal Matching OpenAI o1 System Card

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.753502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.753502Z digest=sha256:bd70d41e62a2309cc9fe47cc7376aaf3b6f9d5a4b7fa5d5f098858089402688e

Observation 503de85d-0ccc-4c78-8803-fd896013e744 · outbound

This paper cites Ya RN : Efficient context window extension of large language models.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Ya RN : Efficient context window extension of large language models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.758028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.758028Z digest=sha256:9e21f0ecc5eb52c6a4b9cc255720138c6b73463453140025ef2ffe843852f5e8

Observation 0111a5d6-db4f-44e7-81b8-c207018617e7 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Roformer: Enhanced transformer with rotary position embedding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.762381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.762381Z digest=sha256:5c26a0e54ee755496e02263f77562ee460e5f50faba186eefeb240f757574559

Observation fad157b5-337f-4d21-a24b-3c112e0c64a0 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.766626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.766626Z digest=sha256:4cafe48819b2fca84aa71d65dc51cdb10e9526a307ed937d643a05d2e66ba059

Observation 4b6c1206-f5ce-43cd-bdaa-a5747401dbd8 · outbound

This paper cites Jamba-1.5: Hybrid Transformer-Mamba Models at Scale.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Jamba-1.5: Hybrid Transformer-Mamba Models at Scale

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.771026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.771026Z digest=sha256:f3e485abbb834a6eed6ed9ff3f414f02b1e067ae6298eb18ac81e95601eef308

Observation ccf53462-1c47-49b6-a4f7-44a843011a39 · outbound

This paper cites Leave no document behind: Benchmarking long-context LLM s with extended multi-doc QA.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Leave no document behind: Benchmarking long-context LLM s with extended multi-doc QA

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.775267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.775267Z digest=sha256:88b02388030c23e3e58ca88555048590296847b0790d21bc922a3429ee9f015b

Observation aaeee15f-17da-422d-b98e-a1520428f6b7 · outbound

This paper cites V., and Zhou, D.

NoLiMa: Long-Context Evaluation Beyond Literal Matching V., and Zhou, D

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T20:07:34.367091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T20:07:33.779724Z digest=sha256:ca36ddb40b42fcb9d7e7b0e3b5e0f1c6fd80e53f132ccdbd6ae32b1134f1a2bf

Observation c9cc6fdd-95be-4b47-81d7-4328c1330e83 · outbound

This paper cites Transformers: State-of-the-art natural language processing.

NoLiMa: Long-Context Evaluation Beyond Literal Matching Transformers: State-of-the-art natural language processing

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.783846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.783846Z digest=sha256:451103e328155fd7335459bff4eaea7d8195ed23752b990541de83ae78a2a9a8

Observation 0118dd4f-d622-4e83-903a-0ef6c525c28f · outbound

This paper cites A., Oguz, B., Khabsa, M., Fang, H., Mehdad, Y., Narang, S., Malik, K., Fan, A., Bhosale, S., Edunov, S., Lewis, M., Wang, S., and Ma, H.

NoLiMa: Long-Context Evaluation Beyond Literal Matching A., Oguz, B., Khabsa, M., Fang, H., Mehdad, Y., Narang, S., Malik, K., Fan, A., Bhosale, S., Edunov, S., Lewis, M., Wang, S., and Ma, H

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.788339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.788339Z digest=sha256:b551159980900fa00d45394ec3cd57a6cd0bc1dfbbe94e2284eb0e0579545691

Observation a61a6ba2-5a9d-4e63-bf3a-bf0dfeedc3fe · outbound

This paper cites HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly.

NoLiMa: Long-Context Evaluation Beyond Literal Matching HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.792487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.792487Z digest=sha256:a605fdb8c75e1cbdc0460e9c6c0448c3213ae7f4f931bfacde01b52236a78034

Observation 613c0194-d071-4128-8a77-b9f1fd881478 · outbound

This paper cites $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens.

NoLiMa: Long-Context Evaluation Beyond Literal Matching $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.797188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.797188Z digest=sha256:5b4f487fa2dbcb27e572b4a913c636954f7fdc4a65bd2edd6c5f65a1e35517f4

Pith citing papers

Observation ba026d5b-c096-4fd6-9b1d-4d898b9f3da7 · inbound

ReadBench: Measuring the Dense Text Visual Reading Ability of Vision-Language Models cites this paper.

ReadBench: Measuring the Dense Text Visual Reading Ability of Vision-Language Models NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:17.501214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:22:17.501214Z digest=sha256:34b6ad7de772e5a9e01c7e010a66c1303f974b48b4788453f0093560cf2144de

Observation 49a3ff40-9ec6-409b-a22a-ef601b0ff7eb · inbound

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions cites this paper.

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T21:20:22.247235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T21:20:22.146417Z digest=sha256:85d0f4785a07897630014aa84b3f4785f1486b303b326983001a2d4b45e2d56f

Observation c3fd6025-7ed3-48ad-89c3-442289975613 · inbound

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions cites this paper.

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T19:36:05.732614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:36:05.732614Z digest=sha256:424062993e97c5f1ba6c7b8a4fe4a84085472230777391f92cd593123a874fec

Observation 3d985bc1-cec7-44e5-919f-670e01084ff8 · inbound

From Alerts to Intelligence: A Novel LLM-Aided Framework for Host-based Intrusion Detection cites this paper.

From Alerts to Intelligence: A Novel LLM-Aided Framework for Host-based Intrusion Detection NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T17:29:59.630413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:29:59.630413Z digest=sha256:1f0f31ad714fa6223f4a660ab6d8eed1ceff8f6443866a0e70bf96a24d92e3f2

Observation 0e1c2769-6ff0-4af4-869d-958632793051 · inbound

Positional Biases Shift as Inputs Approach Context Window Limits cites this paper.

Positional Biases Shift as Inputs Approach Context Window Limits NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T22:12:56.230603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:12:56.230603Z digest=sha256:e37b5f731fa1eb981fea7424db48ee859cb361c583a3e93eef5f9564d10943b7

Observation 7fa9bf73-431d-44e8-9143-c25062bc943f · inbound

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs cites this paper.

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T12:42:31.510188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:42:31.510188Z digest=sha256:10273763566558e7963ec42b3b29d1f2d3d1c22d20e885c731384269562cfcf5

Observation f8cfe597-103c-4256-9874-a5fd89f8bcdb · inbound

MiMo-V2-Flash Technical Report cites this paper.

MiMo-V2-Flash Technical Report NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:33:32.850729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T11:33:32.568261Z digest=sha256:ebe37f0dd0ef3d56b246455645747d12809c87b858105cc1d6e4182a748bcdff

Observation ebb45a47-5775-4c01-b8c5-d78baffdbaa6 · inbound

MemOCR: Layout-Aware Visual Memory for Efficient Long-Horizon Reasoning cites this paper.

MemOCR: Layout-Aware Visual Memory for Efficient Long-Horizon Reasoning NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-21T15:20:17.559167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T15:15:22.055616Z digest=sha256:e7ac09596c564629e94e6e562b733a6c2d017a64471f7855f576a8f5a1929c7b

Observation 822a91a7-5d57-4656-8c1d-38e28f447256 · inbound

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence cites this paper.

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T11:29:58.743708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T11:28:29.772341Z digest=sha256:4b1d37e13bc7e0577885cecdd361759dffcbf9214f2f0443b400e2d013ba94f1

Observation c5068b9f-7f3f-4881-9c34-727a206c78a0 · inbound

Context Collapse: Barriers to Adoption for Generative AI in Workplace Settings cites this paper.

Context Collapse: Barriers to Adoption for Generative AI in Workplace Settings NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:35:52.613651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:58:18.436603Z digest=sha256:7d1d83df5b2129406dcb9ece7ada72ccd76fd7cad402e3e934cf460a6c85c572

Observation b8927049-1f1f-4337-b37f-0f8e46c533fb · inbound

Tool Attention Is All You Need: Dynamic Tool Gating and Lazy Schema Loading for Eliminating the MCP/Tools Tax in Scalable Agentic Workflows cites this paper.

Tool Attention Is All You Need: Dynamic Tool Gating and Lazy Schema Loading for Eliminating the MCP/Tools Tax in Scalable Agentic Workflows NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:31:05.398766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T21:44:15.689372Z digest=sha256:8c1b7df012a499c07bcb518cbe037fa2d9d3adf000a83ae5b0f4295f6cea89b4

Observation da8a7ff6-c924-4d3b-8f91-96429dcce240 · inbound

Retrieval from Within: An Intrinsic Capability of Attention-Based Models cites this paper.

Retrieval from Within: An Intrinsic Capability of Attention-Based Models NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:10.616551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T14:44:21.325977Z digest=sha256:8cce104f4fc183a69e941bbc78ac1d6821192ef9a372264487ea871cc79b953c

Observation eea11f69-7032-4e69-a4f3-53fd3763f342 · inbound

Retrieval from Within: An Intrinsic Capability of Attention-Based Models cites this paper.

Retrieval from Within: An Intrinsic Capability of Attention-Based Models NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:50:55.374291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T01:03:11.473175Z digest=sha256:ccef9eaafc44d817f84e1357a6f37e0b223d44342fd7a7465a2a825e9df6f2e0

Observation afa493ee-83e4-488e-9a27-75874489c4c6 · inbound

Slipstream: Trajectory-Grounded Compaction Validation for Long-Horizon Agents cites this paper.

Slipstream: Trajectory-Grounded Compaction Validation for Long-Horizon Agents NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:06:13.421804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T01:04:55.618417Z digest=sha256:6498cd770d11b36f367bb61a4ebb02e720e9c99097cfcee597538be1be2b4949

Observation 01fa599f-4bba-477e-b716-7d662e871328 · inbound

Where Does Long-Context Supervision Actually Go? Effective-Context Exposure Balancing cites this paper.

Where Does Long-Context Supervision Actually Go? Effective-Context Exposure Balancing NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:51:29.109575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:52:45.320454Z digest=sha256:287741aaae4583e12245c06faeafe38f115d0a9a27185b5b07acd582a49e3d61

Observation 96bd3231-b603-4dce-a24b-ac1144e570f4 · inbound

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks cites this paper.

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:00:21.860068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-25T04:58:15.184063Z digest=sha256:defd965c4520d52a854e4a390195fda9cc7e8a08f2e1ba27a6616eb1332e367b

Observation 03381e0a-2bbe-48e6-b3c7-f1bcca845d58 · inbound

Parallel Context Compaction for Long-Horizon LLM Agent Serving cites this paper.

Parallel Context Compaction for Long-Horizon LLM Agent Serving NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-25T04:36:36.608423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T04:36:22.084650Z digest=sha256:f0b3ce996de7843eac7d464fef4d57f3e3ae0b38be7644ff3a235fd2a23896ef

Observation e680f4fa-dea6-4634-b40a-2ea29b31ef26 · inbound

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions cites this paper.

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:34:48.486069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T15:17:37.904831Z digest=sha256:11c44d4be6b33dfe8cfa2e8cafa0a465f4a04083c1aa24a4760230dc19e86cbf

Observation 75280bb5-728d-4dfd-983b-4d1b7922e2f2 · inbound

Separating Semantic Competition from Context Length in RAG Reading cites this paper.

Separating Semantic Competition from Context Length in RAG Reading NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:03:48.173672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T17:56:05.596190Z digest=sha256:05f311db7382d486336f253ae43330e7a36e29bda18350c88006b7a26862d1e5

Observation 5a26e4c8-dbbd-4d53-ad38-093d68e8730d · inbound

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators cites this paper.

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:19.684811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T15:02:47.349477Z digest=sha256:ab5deb26066edd9e847e61346b4f7e089d9b5e51452450df7c3a401dd71cd981

Observation a2501fa4-2a6a-4826-901b-c85e3d308ca2 · inbound

LifeSide: Benchmarking Agents as Lifelong Digital Companions cites this paper.

LifeSide: Benchmarking Agents as Lifelong Digital Companions NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:46:46.116061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T06:42:55.108719Z digest=sha256:875ae56996566c7e23a65b4ff404a751394d01c07b526243deca8e8b2d460f8b

Observation aa527bb7-cc01-4766-9226-99b29e52e80c · inbound

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces cites this paper.

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 154

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:31:21.412334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T22:22:52.690010Z digest=sha256:26def17a5abb5de0c2016c14c8f18048650068fb7e7fcf550b313c4fa7f3fe11

Observation 61b06e64-cbc2-4a9a-aa87-b5e910571d39 · inbound

EASE-TTT: Evidence-Aligned Selective Test-Time Training for Long-Context Question Answering cites this paper.

EASE-TTT: Evidence-Aligned Selective Test-Time Training for Long-Context Question Answering NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:17:15.125201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T22:05:00.537690Z digest=sha256:8f11173f2112dc16474e99a8d1438ea2155852f538648d1e499d807959c51f74

Observation 3006c052-3bf8-4124-aff9-f5bb0270388e · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 82

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:43.995595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:30b8222f250bfd9a150259cf401777179de205f96a90a774f55ca44c33fe327c

Observation 7aaa31c2-7101-41a3-954f-e04bd9740c6b · inbound

UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp cites this paper.

UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-14T10:51:16.019022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:51:16.019022Z digest=sha256:0665a95964bcc869d1ace91370e982bbc91aedc3c59e161f5dc73686d49d7923

Observation 6be85427-c3b4-414b-8998-b3301b7383e5 · inbound

Where Facts Go Missing: A Layerwise Taxonomy and Per-Layer Attribution of Information Omission in Air-Gapped LLMAgent Pipelines cites this paper.

Where Facts Go Missing: A Layerwise Taxonomy and Per-Layer Attribution of Information Omission in Air-Gapped LLMAgent Pipelines NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T04:48:27.492158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:48:27.492158Z digest=sha256:5c1eaf7143f1f6e3dd067954a8dfea6ab5a08649020a4a0d15caa94b9c974f90

Observation 144aee59-b3cf-4cbf-b526-4207cec02b4e · inbound

CORVUS: Context Optimization and Reduction Via Underlying Synchronization for LLM Coding Agents cites this paper.

CORVUS: Context Optimization and Reduction Via Underlying Synchronization for LLM Coding Agents NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T15:05:51.135681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:05:51.135681Z digest=sha256:26379e882c630f79bf99c3809684b4a4b2e68d586cf70d1ea5c79d3b5dd8c512

Observation 1428011f-2a66-4d73-8bba-2b7221fe30af · inbound

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following cites this paper.

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following NoLiMa: Long-Context Evaluation Beyond Literal Matching

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:36:08.596712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:36:08.596712Z digest=sha256:3b168e26a1e22f3e099558fba1b952f55e1a81dfa2a74e51e5620fb28a0d4951