Pith. sign in

Paper Citation Record · LEDGER

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory

As of 14 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2607.27919.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.27919 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-31T23:07:18.588999Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved40
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f3335169-b4d0-4f8c-928d-071e91f185e7 · outbound

This paper cites an unresolved cited work.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.237143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.237143Z digest=sha256:db2bf4c9a5e2edb0bd34c02c3ebba19d5693c5270968d468798991c7c0938979

Observation f33655f4-5c94-4c79-b548-75e8df3eceb1 · outbound

This paper cites Output: number EC4.3.3.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Output: number EC4.3.3

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.326633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.326633Z digest=sha256:4000958c4448b5445f54dc740ff0388ddd28cf4711300d05638de6f3c53f9f49

Observation afc6430e-0aa1-4efd-8cfb-60f769502f09 · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.027807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.027807Z digest=sha256:96fdad1d86f6a28c9e230b6bf17eafbe1dd16a0c2302699dc3aa755c903ae1c3

Observation d6e2000b-bdf5-45a0-a693-8ab7eeb68f05 · outbound

This paper cites an unresolved cited work.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.222219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.222219Z digest=sha256:624148651c453ac7cdce536e052f9e92b4a4e81de8f12981b2a973b415a83ba7

Observation d1783bd4-9bb7-4f53-8616-0f3f9c253f4a · outbound

This paper cites Biology-instructions: A dataset and benchmark for multi-omics sequence understanding capability of large language models.arXiv preprint arXiv:2412.19191,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Biology-instructions: A dataset and benchmark for multi-omics sequence understanding capability of large language models.arXiv preprint arXiv:2412.19191,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.040894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.040894Z digest=sha256:87758398f99943a7c52bf6421a846bc52bdc3dd048a80e463a0a081e2f2f32f4

Observation 341b64aa-0496-43df-82d8-94047934b1e7 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Measuring Massive Multitask Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.044920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.044920Z digest=sha256:5d774bd9a5cb9e9bb074d1a9ea3362f539e3dab40cf1d885cd4a2d30dcc0e44c

Observation 423d9f94-45d4-4634-9169-fa4252f912f6 · outbound

This paper cites Generalization through Memorization: Nearest Neighbor Language Models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Generalization through Memorization: Nearest Neighbor Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.054189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.054189Z digest=sha256:440025c7b30ee140dccbc9ad02be4a2a1b80295a144e941af8b4e6a661f6cd11

Observation 96cf3753-c46c-465d-90c0-498e0e1d7dc1 · outbound

This paper cites Halueval: A large-scale halluci- nation evaluation benchmark for large language models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Halueval: A large-scale halluci- nation evaluation benchmark for large language models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.058658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.058658Z digest=sha256:fd9dbea56c7de41c25095810b211335eab832db71cd268bf3595c7ce23fca5a7

Observation 2cbe41f1-994d-42c5-b755-0e9569b9af4b · outbound

This paper cites LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.062768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.062768Z digest=sha256:2ce17eb940d1eebc81dd467efb1fe89c2614fbb97a21518581229534d9184aa3

Observation d7db37db-285e-4641-9eec-daa7e2f97f4d · outbound

This paper cites Decoupled Weight Decay Regularization.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Decoupled Weight Decay Regularization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.066706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.066706Z digest=sha256:ba3cfb48d62e10d6fde06c29f42ca1eafa381e51fcafbc9230948d58f365323e

Observation 378ba405-6464-41b8-8757-824308aa3560 · outbound

This paper cites Olmo 3.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Olmo 3

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.075283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.075283Z digest=sha256:f90340c560c1b2e25006cae5922c25621962047321e31daef75233e1ae0091d1

Observation 35be0422-3129-40e8-b6c8-dfae8f463175 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.083883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.083883Z digest=sha256:d357699ef0a2b524529e2489c993f5b5b6e21d8614e86a3afe7cf19d044100df

Observation b7d365a7-83a0-4e9a-b5c0-426545b4631a · outbound

This paper cites OpenAI GPT-5 System Card.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory OpenAI GPT-5 System Card

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.092517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.092517Z digest=sha256:3ad8b5bce5c34d53b469d0581c9b4258ab6a6594535cfcf64a6553d59e11cf42

Observation 11195077-0009-4f00-bf43-05f7944218fb · outbound

This paper cites Memoryandbrainsystems: 1969–2009.JournalofNeuroscience,29(41):12711–12716,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Memoryandbrainsystems: 1969–2009.JournalofNeuroscience,29(41):12711–12716,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.096594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.096594Z digest=sha256:4b4d01f90696e3afe9273c0bfd333872e3a3160d54b24a258f3422390066dc24

Observation 52a03c36-fd7e-4d10-b040-e977c38cc231 · outbound

This paper cites Qwen3.5-Omni Technical Report.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Qwen3.5-Omni Technical Report

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.100553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.100553Z digest=sha256:11c99a7aac5176ada360c7bb303cc9c3bc5d9bb1dd25692461a93548b24d2dbb

Observation e5e3f4c1-ba32-4d5f-bf23-320064905b76 · outbound

This paper cites Infllm: Training-free long-context extrapolation for llms with an efficient context memory.Advances in neural information processing systems, 37:119638–119661, 2024a.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Infllm: Training-free long-context extrapolation for llms with an efficient context memory.Advances in neural information processing systems, 37:119638–119661, 2024a

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.108314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.108314Z digest=sha256:485839cb569f9d00ee31b1751d9f899898f0cbac4d65c07ea942d0781568856f

Observation 22eb6e8b-7c22-4202-af09-2c36777c9400 · outbound

This paper cites Hotpotqa: A dataset for diverse, explainable multi-hop question answering.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Hotpotqa: A dataset for diverse, explainable multi-hop question answering

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.116807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.116807Z digest=sha256:685d27366f7045121e3c393415087fa66bc077fa796f66ad18c8156b8f997b64

Observation b15debcf-38a5-438b-8e56-588ce5b5a5b0 · outbound

This paper cites Vismem: Latent vision memory unlocks potential of vision-language models.arXiv preprint arXiv:2511.11007,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Vismem: Latent vision memory unlocks potential of vision-language models.arXiv preprint arXiv:2511.11007,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.121587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.121587Z digest=sha256:8439068d8e4ab57c4c2f809eeff4fbccf241d65468691d18d6ed6080e5619394

Observation 83b901b2-45f7-4d04-86cf-772714be97c9 · outbound

This paper cites DISC-LawLLM: Fine-tuning Large Language Models for Intelligent Legal Services.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory DISC-LawLLM: Fine-tuning Large Language Models for Intelligent Legal Services

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.133745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.133745Z digest=sha256:ee91b4d8c22274f615e8abe72feb8a1db60dba1e81492102e3dda1eba90ac0c7

Observation ca6e958a-2c2b-4beb-9635-29baa4eaf2c9 · outbound

This paper cites GLM-5: from Vibe Coding to Agentic Engineering.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory GLM-5: from Vibe Coding to Agentic Engineering

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.155008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.155008Z digest=sha256:71d8fe253ee1df53cb2d2d82b2e2c355325ea951524d0c2050246d2f17314326

Observation f558500c-4181-4514-a2d8-e928c771ae81 · outbound

This paper cites Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.172943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.172943Z digest=sha256:543227f4fd6d74cef7f09c7a96cb7731888fca9c82812eb6abebbf1a72101ac2

Observation 7c458226-6e37-426a-a863-7f406850b22b · outbound

This paper cites Pre-training limited memory language models with internal and external knowledge.arXiv preprint arXiv:2505.15962,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Pre-training limited memory language models with internal and external knowledge.arXiv preprint arXiv:2505.15962,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.193327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.193327Z digest=sha256:d8dc9320b73b56046584cca622eea81c8cd4275140ba95e34607eacd8ae57a49

Observation a7dc31d8-0c1c-406d-b7ab-0beee8b9ef3c · outbound

This paper cites Training language models with memory augmentation.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Training language models with memory augmentation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.205826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.205826Z digest=sha256:202344df09edf37c6e86e388254d53d1d0ee78cee721c5380eb69fe681293692

Observation 265d62b5-671d-4aa4-9d89-ae5fe79bebb8 · outbound

This paper cites Table 14| Top-3 versus top-5 retrieval for the RAG baseline.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Table 14| Top-3 versus top-5 retrieval for the RAG baseline

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.255922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.255922Z digest=sha256:1ab879335090b2b5b731c62e6f8417164c28e12a1b7872485cb0ebf996e2b0e8

Observation 0b08c9d3-ac7e-4057-92d5-8fba4e260753 · outbound

This paper cites Zaheer Khan, Yuvraj Singh and Marlon Samuels made their ODI debuts during the competition.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Zaheer Khan, Yuvraj Singh and Marlon Samuels made their ODI debuts during the competition

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.276958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.276958Z digest=sha256:d1a0010635059188cece7b685ba0b3075150a64df3a8b91975bff410a69c0ce5

Observation f28af8c7-e7f6-48a4-a693-a5666f186495 · outbound

This paper cites BioInst Table 21|Per-task domain memory results on BioInst for the Qwen3-0.6B-Base backbone.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory BioInst Table 21|Per-task domain memory results on BioInst for the Qwen3-0.6B-Base backbone

Reference 42

Resolution
malformed identifier
no resolver link, observed 2026-07-31T23:07:18.434809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.434809Z digest=sha256:80d6205dc7bb240e25766d77a75b0179d5d41d8dd4d346e176f5d5678902db0b

Observation f33e5e3d-d57a-4d2d-8148-3737ab225eb3 · outbound

This paper cites AVG is the macro average over the 25 task scores.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory AVG is the macro average over the 25 task scores

Reference 100

Resolution
malformed identifier
no resolver link, observed 2026-07-31T23:07:18.588999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.588999Z digest=sha256:162c73e4e951226f8a8b5e29936f052e8adf4108fbc5b5cd7d79190f37535bca

Observation cb505e47-680a-46a1-9f1c-d4656644d17e · outbound

This paper cites Lawbench: Benchmarking legal knowledge of large language models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Lawbench: Benchmarking legal knowledge of large language models

Reference 381

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.023589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.023589Z digest=sha256:b7341eabd6c7a55a4a556561bb841c17bd8f1cdabb586822ad9b2b1643b71ca0

Observation 8b294cf3-c5d9-4b6b-9a7b-252e767cc263 · outbound

This paper cites URL https://doi.org/10.1371/journal.pcbi.100.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory URL https://doi.org/10.1371/journal.pcbi.100

Reference 2009

Resolution
verified exact
doi, observed 2026-07-31T23:11:01.454730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-31T23:07:18.018810Z digest=sha256:e0c38ffa5eb30de7d5f293ddafe00564eb0ae86c93130c66019d780e03d5d475

Observation 750f4510-084a-45a0-b847-f0ab59c62d15 · outbound

This paper cites Gimme Shelter.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Gimme Shelter

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.300997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.300997Z digest=sha256:150200ea4f15001e8fc9d4130017b0cdc94cecbb782828cd5eb7d91343c37b26

Observation 6f7cf464-a07f-4cb2-afde-046a4380e2c0 · outbound

This paper cites Deepsieve: Information sieving via llm-as-a-knowledge-router.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Deepsieve: Information sieving via llm-as-a-knowledge-router

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.032647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.032647Z digest=sha256:5f022a3a2105de720c42fa492d710c6205125e7ce33bc702f1f44137a74b97a0

Observation 10935bf2-8bd0-4765-8873-8ba4feda3d39 · outbound

This paper cites Mlp mem- ory: A retriever-pretrained memory for large language models.arXiv preprint arXiv:2508.01832,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Mlp mem- ory: A retriever-pretrained memory for large language models.arXiv preprint arXiv:2508.01832,

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.104442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.104442Z digest=sha256:c804b87ba9995b35a7a188914517ef719ddf59b1b2cbebf2c951bc4c8cb8da37

Observation 0f687444-f0ab-4b86-b9a6-2b67db7e77ae · outbound

This paper cites Measuring and narrowing the compositionality gap in language models.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Measuring and narrowing the compositionality gap in language models

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.080021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.080021Z digest=sha256:53abe0a7d7e038bb0514496e625c3580c8742dd310a15fb03c74680481d892cc

Observation c13c17a5-33fd-4537-9958-6f3cf84b708a · outbound

This paper cites Demystifying domain- adaptive post-training for financial llms.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Demystifying domain- adaptive post-training for financial llms

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.049255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.049255Z digest=sha256:408552d3d77ad7edd035bb74f90650e3e6b49c107f5bfac6fdfcff673601e206

Observation 9ded397c-b5b0-4595-8dd4-ce7610ec1ad0 · outbound

This paper cites Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.005372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.005372Z digest=sha256:ca80c0883e1d5976fcdd6bfa92fe53c1fbc7dd61a4be288462a4b8a0adfad9d3

Observation 0825969d-43f8-4d20-bd03-8d857dd9c58e · outbound

This paper cites LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.010236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.010236Z digest=sha256:88ca919e7c6b6ff68735defce09178bb6d4b1eb718254d41fb27e9db344ef52e

Observation 049b80c2-e55b-45a9-86c7-d377d65bc0c6 · outbound

This paper cites Measuring memorization in language models via probabilistic extraction.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Measuring memorization in language models via probabilistic extraction

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.036737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.036737Z digest=sha256:b3e8750cd57c7801ff155556c70aee2be1d63e419816791ffc860fdcbdacb02b

Observation c8d0ba66-3540-460f-a7d9-5e957dc98f8a · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.088076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.088076Z digest=sha256:a2dd436dab2cb2c1402cfa9dd8a67148e7505b93e97ce6dd4e791d7e4a9ba431

Observation 6ea5ff70-16ff-4f12-88f1-714f37dbe741 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.000927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.000927Z digest=sha256:0b6a4a58f439523634f0214255ba11ce913005c335309aca3cc8cdb22b0d0b37

Observation 4b3af424-dd33-4748-91ab-817261791c7a · outbound

This paper cites 2 OLMo 2 Furious.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory 2 OLMo 2 Furious

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.071026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.071026Z digest=sha256:6046d623bf27fcd11b66fe8f697e30bfab6926a54126ed9e61d67b3f2a7f363a

Observation f26b7e36-ce8a-4fed-8507-96f858ff369a · outbound

This paper cites Titans: Learning to Memorize at Test Time.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Titans: Learning to Memorize at Test Time

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:17.996097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:17.996097Z digest=sha256:86ac376fce8cab5555225e69a88c2d8a277700306add663aed9595de6e92c455

Observation f5233451-7b0a-4c15-871a-32b219e719f7 · outbound

This paper cites Twinvoice: A multi-dimensional benchmark towards digital twins via llm persona simulation.arXiv preprint arXiv:2510.25536,.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Twinvoice: A multi-dimensional benchmark towards digital twins via llm persona simulation.arXiv preprint arXiv:2510.25536,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.014654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.014654Z digest=sha256:ca3a69dbd9e80da6cdd20b5e6a34f0976dc9ce55e5f0bf91923c15ed95ca4bcd

Observation e4e18648-9f16-4944-b8bc-3051f0351567 · outbound

This paper cites Qwen3 Technical Report.

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Qwen3 Technical Report

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:18.112592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:07:18.112592Z digest=sha256:19554d4f15a009d4eb3f38c48a2f03ba8d4a806f8694f5dd549683abbb46e63e

Pith citing papers

No inbound Pith citation observations are available.