Pith. sign in

Paper Citation Record · LEDGER

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory

As of 7 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2607.27919.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.27919 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-31T23:07:18.588999Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved40
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f3335169-b4d0-4f8c-928d-071e91f185e7 · outbound

This paper cites an unresolved cited work.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.237143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.237143Z digest=sha256:844a11e81419d598d2df99496eff7543e8ca440fb0d972b8a44f0e1bd9b5efa1

Observation f33655f4-5c94-4c79-b548-75e8df3eceb1 · outbound

This paper cites Output: number EC4.3.3.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Output: number EC4.3.3

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.326633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.326633Z digest=sha256:69f9f4631e57f60c1fba6ed20f67844b7b74650158d786b94497c806a3ce6d14

Observation afc6430e-0aa1-4efd-8cfb-60f769502f09 · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.027807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.027807Z digest=sha256:78276df7420b10fc3b292020500729d3ef1900a781db2bb02a7bf0d84037f25c

Observation d6e2000b-bdf5-45a0-a693-8ab7eeb68f05 · outbound

This paper cites an unresolved cited work.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.222219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.222219Z digest=sha256:e35594327cc8690596b7fd7847ac60005280d8f75d3aac69c76b4708d8376649

Observation d1783bd4-9bb7-4f53-8616-0f3f9c253f4a · outbound

This paper cites Biology-instructions: A dataset and benchmark for multi-omics sequence understanding capability of large language models.arXiv preprint arXiv:2412.19191,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Biology-instructions: A dataset and benchmark for multi-omics sequence understanding capability of large language models.arXiv preprint arXiv:2412.19191,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.040894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.040894Z digest=sha256:bf95b6c9ca71413694aa4abd0387959a2247e3117aec6134274475a56daeaf29

Observation 341b64aa-0496-43df-82d8-94047934b1e7 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Measuring Massive Multitask Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.044920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.044920Z digest=sha256:4415c306518b3700d5360aaddc3e08f3fab9ed6d02c9d4aec805305c5b4072e0

Observation 423d9f94-45d4-4634-9169-fa4252f912f6 · outbound

This paper cites Generalization through Memorization: Nearest Neighbor Language Models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Generalization through Memorization: Nearest Neighbor Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.054189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.054189Z digest=sha256:4f9ab2aba502cee056fa24a8fee075f72c9856e1a621410c6b25c5b0a0c1019c

Observation 96cf3753-c46c-465d-90c0-498e0e1d7dc1 · outbound

This paper cites Halueval: A large-scale halluci- nation evaluation benchmark for large language models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Halueval: A large-scale halluci- nation evaluation benchmark for large language models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.058658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.058658Z digest=sha256:2ca4fa136c43fc14b580436a72b7128bd44d17548fd0b688739fc6800da66510

Observation 2cbe41f1-994d-42c5-b755-0e9569b9af4b · outbound

This paper cites LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.062768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.062768Z digest=sha256:7df2d799690de7c780c14f25cafb71c8283483ecc980cbb19df500cec348f8c2

Observation d7db37db-285e-4641-9eec-daa7e2f97f4d · outbound

This paper cites Decoupled Weight Decay Regularization.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Decoupled Weight Decay Regularization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.066706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.066706Z digest=sha256:12f3c891999b4196a028a16b2a07623bbd688273bcbf142645bc77d41c888fa5

Observation 378ba405-6464-41b8-8757-824308aa3560 · outbound

This paper cites Olmo 3.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Olmo 3

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.075283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.075283Z digest=sha256:389c072ace06792e58ff9331c8aecefe4981ad14fc069015b0d8df186c9812ec

Observation 35be0422-3129-40e8-b6c8-dfae8f463175 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.083883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.083883Z digest=sha256:bafd136272183a365423715f17e55655e61e4629962ec739ac1f6ccc7f605610

Observation b7d365a7-83a0-4e9a-b5c0-426545b4631a · outbound

This paper cites OpenAI GPT-5 System Card.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory OpenAI GPT-5 System Card

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.092517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.092517Z digest=sha256:b29a2955b0d1d38ded63526a9b86b108f9c7e4ef38c9f4b2f771c4f4f32781d1

Observation 11195077-0009-4f00-bf43-05f7944218fb · outbound

This paper cites Memoryandbrainsystems: 1969–2009.JournalofNeuroscience,29(41):12711–12716,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Memoryandbrainsystems: 1969–2009.JournalofNeuroscience,29(41):12711–12716,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.096594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.096594Z digest=sha256:073dceb03b32473a952c36c391785446eb3fe43d8586233d36657457e00ab53a

Observation 52a03c36-fd7e-4d10-b040-e977c38cc231 · outbound

This paper cites Qwen3.5-Omni Technical Report.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Qwen3.5-Omni Technical Report

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.100553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.100553Z digest=sha256:b331d8d44735a39fcc56577002895b0b11b74c4e15008713442819eb706c7dd8

Observation e5e3f4c1-ba32-4d5f-bf23-320064905b76 · outbound

This paper cites Infllm: Training-free long-context extrapolation for llms with an efficient context memory.Advances in neural information processing systems, 37:119638–119661, 2024a.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Infllm: Training-free long-context extrapolation for llms with an efficient context memory.Advances in neural information processing systems, 37:119638–119661, 2024a

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.108314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.108314Z digest=sha256:79b5ce69413ff5a5f9c599bf0bd240cb2e1f3c5e6a99229ba0d3ceb91e76be7d

Observation 22eb6e8b-7c22-4202-af09-2c36777c9400 · outbound

This paper cites Hotpotqa: A dataset for diverse, explainable multi-hop question answering.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Hotpotqa: A dataset for diverse, explainable multi-hop question answering

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.116807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.116807Z digest=sha256:5c356d663bebd136c0547cd6c8b19aff37d267ec70d82a6754e6a6aab6e13d85

Observation b15debcf-38a5-438b-8e56-588ce5b5a5b0 · outbound

This paper cites Vismem: Latent vision memory unlocks potential of vision-language models.arXiv preprint arXiv:2511.11007,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Vismem: Latent vision memory unlocks potential of vision-language models.arXiv preprint arXiv:2511.11007,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.121587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.121587Z digest=sha256:fe082faee4cd125c14465df5071d1eae7ef4c8aa534e0fd04f047c16e5555363

Observation 83b901b2-45f7-4d04-86cf-772714be97c9 · outbound

This paper cites DISC-LawLLM: Fine-tuning Large Language Models for Intelligent Legal Services.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory DISC-LawLLM: Fine-tuning Large Language Models for Intelligent Legal Services

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.133745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.133745Z digest=sha256:da5adad1b617ec3e7e34b0f2b0154a44b88cbe235ef7f0c7c2a01b22b657fde0

Observation ca6e958a-2c2b-4beb-9635-29baa4eaf2c9 · outbound

This paper cites GLM-5: from Vibe Coding to Agentic Engineering.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory GLM-5: from Vibe Coding to Agentic Engineering

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.155008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.155008Z digest=sha256:4e5238b4652e8cdceeec058e911a775e16811a189d9fd846569592063bfec334

Observation f558500c-4181-4514-a2d8-e928c771ae81 · outbound

This paper cites Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.172943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.172943Z digest=sha256:e96be4ba15bb1a030d147cd0f5e3dc95ac4d51cc328d5d8bb00b85c4236cf9a3

Observation 7c458226-6e37-426a-a863-7f406850b22b · outbound

This paper cites Pre-training limited memory language models with internal and external knowledge.arXiv preprint arXiv:2505.15962,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Pre-training limited memory language models with internal and external knowledge.arXiv preprint arXiv:2505.15962,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.193327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.193327Z digest=sha256:a0d90d2a7287362e65c316420bc3fbcaca86d5636dad8ede5d9ace2cbda47b8e

Observation a7dc31d8-0c1c-406d-b7ab-0beee8b9ef3c · outbound

This paper cites Training language models with memory augmentation.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Training language models with memory augmentation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.205826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.205826Z digest=sha256:bd9344f9dda8bde0c7fd7ff3cda0a15be30f6287ee644c5f5c8ce9fbf3f6b131

Observation 265d62b5-671d-4aa4-9d89-ae5fe79bebb8 · outbound

This paper cites Table 14| Top-3 versus top-5 retrieval for the RAG baseline.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Table 14| Top-3 versus top-5 retrieval for the RAG baseline

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.255922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.255922Z digest=sha256:307049bc8ec02291562b6ed30257a22ab8bfc712b5c2bd1853dfc976ab5a77e7

Observation 0b08c9d3-ac7e-4057-92d5-8fba4e260753 · outbound

This paper cites Zaheer Khan, Yuvraj Singh and Marlon Samuels made their ODI debuts during the competition.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Zaheer Khan, Yuvraj Singh and Marlon Samuels made their ODI debuts during the competition

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.276958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.276958Z digest=sha256:72f06e801da859e4c3c3111a80aaa865c7af09011e021043b7b24b3dc94ed4b6

Observation f28af8c7-e7f6-48a4-a693-a5666f186495 · outbound

This paper cites BioInst Table 21|Per-task domain memory results on BioInst for the Qwen3-0.6B-Base backbone.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory BioInst Table 21|Per-task domain memory results on BioInst for the Qwen3-0.6B-Base backbone

Reference 42

Resolution
malformed identifier
no resolver link, observed 2026-07-31T23:07:18.434809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.434809Z digest=sha256:6c7c26a1f7ff84e56b794a40fb83610984c6d580204ac7183b0446c003ef9665

Observation f33e5e3d-d57a-4d2d-8148-3737ab225eb3 · outbound

This paper cites AVG is the macro average over the 25 task scores.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory AVG is the macro average over the 25 task scores

Reference 100

Resolution
malformed identifier
no resolver link, observed 2026-07-31T23:07:18.588999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.588999Z digest=sha256:3d648379b2b32424c5ccaac0b50b147435ebd2f736c0fdc265c55e943a1394d2

Observation cb505e47-680a-46a1-9f1c-d4656644d17e · outbound

This paper cites Lawbench: Benchmarking legal knowledge of large language models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Lawbench: Benchmarking legal knowledge of large language models

Reference 381

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.023589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.023589Z digest=sha256:c17884e9b09694a784bec589847e9753d2de781e812d16d705a499d7b6724c95

Observation 8b294cf3-c5d9-4b6b-9a7b-252e767cc263 · outbound

This paper cites URL https://doi.org/10.1371/journal.pcbi.100.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory URL https://doi.org/10.1371/journal.pcbi.100

Reference 2009

Resolution
verified exact
doi, observed 2026-07-31T23:11:01.454730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-31T23:07:18.018810Z digest=sha256:1bd03743f89468e391f6fa514b31067059244de0b45b6ba160654dbceff3bf94

Observation 750f4510-084a-45a0-b847-f0ab59c62d15 · outbound

This paper cites Gimme Shelter.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Gimme Shelter

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.300997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.300997Z digest=sha256:5bceda841577b1455c31706f88aa958183c76accad8367addf6e250a0a2ad527

Observation 6f7cf464-a07f-4cb2-afde-046a4380e2c0 · outbound

This paper cites Deepsieve: Information sieving via llm-as-a-knowledge-router.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Deepsieve: Information sieving via llm-as-a-knowledge-router

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.032647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.032647Z digest=sha256:45c083cfefb56587d6775459f1d1521a01e8fdab41be7ebc9dc8292721d1813c

Observation 10935bf2-8bd0-4765-8873-8ba4feda3d39 · outbound

This paper cites Mlp mem- ory: A retriever-pretrained memory for large language models.arXiv preprint arXiv:2508.01832,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Mlp mem- ory: A retriever-pretrained memory for large language models.arXiv preprint arXiv:2508.01832,

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.104442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.104442Z digest=sha256:5cb3efd20ad13a12f7e55f5f502d71f1c7d5d14e8c24e91b4ad6fd0ebc1c79cd

Observation 0f687444-f0ab-4b86-b9a6-2b67db7e77ae · outbound

This paper cites Measuring and narrowing the compositionality gap in language models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Measuring and narrowing the compositionality gap in language models

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.080021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.080021Z digest=sha256:736ef296ac216e9fa2017ca760febc96a35cabc258e5fefe0e1976c94f97680c

Observation c13c17a5-33fd-4537-9958-6f3cf84b708a · outbound

This paper cites Demystifying domain- adaptive post-training for financial llms.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Demystifying domain- adaptive post-training for financial llms

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.049255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.049255Z digest=sha256:6b302d791ae459a8cd1675678945b4270fe61fd24ce3ed87e77cddb292f80a48

Observation 9ded397c-b5b0-4595-8dd4-ce7610ec1ad0 · outbound

This paper cites Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.005372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.005372Z digest=sha256:fd74563033a07d872bd49aebc8458f494c0e1f568fcd0953ad54f690254ff18f

Observation 0825969d-43f8-4d20-bd03-8d857dd9c58e · outbound

This paper cites LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.010236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.010236Z digest=sha256:8f85fa12783a02d8b53387b5ec624b711c89f760bd87b58578a70071b89185c5

Observation 049b80c2-e55b-45a9-86c7-d377d65bc0c6 · outbound

This paper cites Measuring memorization in language models via probabilistic extraction.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Measuring memorization in language models via probabilistic extraction

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.036737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.036737Z digest=sha256:5c7f2e926d69b1ffa25c952db42578c51ac6b3805849adf6243667df111d7782

Observation c8d0ba66-3540-460f-a7d9-5e957dc98f8a · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.088076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.088076Z digest=sha256:abf4f87c2518f9563db8feff50623436bb4d7161a058b714afbf99402941fbe6

Observation 6ea5ff70-16ff-4f12-88f1-714f37dbe741 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.000927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.000927Z digest=sha256:2e63bc9002fcdf43553261ac512982e3e892baafb6085cb53403eab532f0b786

Observation 4b3af424-dd33-4748-91ab-817261791c7a · outbound

This paper cites 2 OLMo 2 Furious.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory 2 OLMo 2 Furious

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.071026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.071026Z digest=sha256:def9da5580ec2477b7cbcbb1c972918910bdb566b3a0e519d5e81f56c8749347

Observation f26b7e36-ce8a-4fed-8507-96f858ff369a · outbound

This paper cites Titans: Learning to Memorize at Test Time.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Titans: Learning to Memorize at Test Time

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:17.996097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:17.996097Z digest=sha256:aa101c7ca17019538ddac804f53fbecaf8cac49947ca9387f23e163158fcaa95

Observation f5233451-7b0a-4c15-871a-32b219e719f7 · outbound

This paper cites Twinvoice: A multi-dimensional benchmark towards digital twins via llm persona simulation.arXiv preprint arXiv:2510.25536,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Twinvoice: A multi-dimensional benchmark towards digital twins via llm persona simulation.arXiv preprint arXiv:2510.25536,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.014654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.014654Z digest=sha256:849b31ad16455458f9f2f7c98deaeb0d23db4a12883e1610b4925e52e25ba96c

Observation e4e18648-9f16-4944-b8bc-3051f0351567 · outbound

This paper cites Qwen3 Technical Report.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Qwen3 Technical Report

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.112592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.112592Z digest=sha256:1af5fc3d6a7055ad378fb9201d09d8a5d66fb1a21b18a89eebc1fde5dc3ef3f9

Pith citing papers

No inbound Pith citation observations are available.