Pith. sign in

Paper Citation Record · LEDGER

TPTT: Transforming Pretrained Transformers into Titans

As of 17 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 2 inbound Pith citation observations for arXiv:2506.17671.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17671 v2

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:34:48.726772Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T14:02:35.245972Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T23:43:51.237854Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1ec4e09b-a51e-4dbf-b4b3-7cade0850fa4 · outbound

This paper cites Qwen Technical Report.

TPTT: Transforming Pretrained Transformers into Titans Qwen Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:45.080735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:45.080735Z digest=sha256:f1ec1285edce53a3cb123bf0c5b6c08c63f95a2321afc9b226e3138ed0fbe664

Observation 8b7ea0d4-e7fd-4d84-b4a8-52538cd4b30a · outbound

This paper cites It's All Connected: A Journey Through Test-Time Memorization, Attentional Bias, Retention, and Online Optimization.

TPTT: Transforming Pretrained Transformers into Titans It's All Connected: A Journey Through Test-Time Memorization, Attentional Bias, Retention, and Online Optimization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:45.116726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:45.116726Z digest=sha256:ac64d970d1406d8547c3a62bdd480345e6398a3c0523fa3ce356a432b391ff30

Observation 34188fda-ff50-4700-a33f-c50609302296 · outbound

This paper cites Titans: Learning to Memorize at Test Time.

TPTT: Transforming Pretrained Transformers into Titans Titans: Learning to Memorize at Test Time

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:45.170666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:45.170666Z digest=sha256:6aa02d121fb346e819afe550158bb14312a93c6e04be28bffcf7c76ee24674d3

Observation daf8618f-a457-4dc5-ba81-59495993e7d3 · outbound

This paper cites Rethinking Attention with Performers.

TPTT: Transforming Pretrained Transformers into Titans Rethinking Attention with Performers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:45.265877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:45.265877Z digest=sha256:a707f31a4fff24260f82297d01eeb73d27fb2ff7faa7f7790b5261a0aae95b14

Observation 47f08d6e-57d1-4a3a-b9a3-2f057d6d172a · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

TPTT: Transforming Pretrained Transformers into Titans FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:45.454745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:45.454745Z digest=sha256:f2249aa9dbf16ebb6ba90d37a7f09ff20b0ff02f1d84537ea72e0b96f3f63ed6

Observation 3572c2d4-deb6-4df9-80a8-02877e4100f9 · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

TPTT: Transforming Pretrained Transformers into Titans Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:45.611500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:45.611500Z digest=sha256:97e615767f1ad613f2752823f2b29780313613bcff5cb766219c6aad1856e439

Observation a5f78074-2cd3-4f7f-a3f5-d5c2d36beb0b · outbound

This paper cites Bert: Pre-training of deep bidi- rectional transformers for language understanding.

TPTT: Transforming Pretrained Transformers into Titans Bert: Pre-training of deep bidi- rectional transformers for language understanding

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:34:51.198763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T23:34:45.698479Z digest=sha256:ec31580aa2191914e48649851babf277528b10bd17524527d088da532aa7d618

Observation b7c92f0a-8261-4796-be48-0e04be0fbe9d · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

TPTT: Transforming Pretrained Transformers into Titans An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:45.800061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:45.800061Z digest=sha256:6224849e8e3b23ee8e7b498fd16bb28bbbadf875c73a399348c33ab08227d7b8

Observation 5d9c5f75-49ee-47f1-ae70-9f44ff6dd367 · outbound

This paper cites Lora - hugging face peft documentation.

TPTT: Transforming Pretrained Transformers into Titans Lora - hugging face peft documentation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:34:51.009732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T23:34:45.897628Z digest=sha256:88080ea65e30192f1052b1a8597f356825a70e4836b871101eda2f1a2a8276b1

Observation 93e42bd2-9816-4616-ac7c-3d3444b57719 · outbound

This paper cites OLMo: Accelerating the Science of Language Models.

TPTT: Transforming Pretrained Transformers into Titans OLMo: Accelerating the Science of Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:46.034748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:46.034748Z digest=sha256:a70a32985951c87958d3b34e23c86d5f7f7cf87323c5f64d665740212f59106f

Observation 4d55c19d-5ad2-4dba-bd2b-ae7be2193eee · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

TPTT: Transforming Pretrained Transformers into Titans Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:46.161689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:46.161689Z digest=sha256:4fb9b069d0a7f9883cd67dba0cf25e28ca3b13ad82c68ecd5719e8f1d2110e4d

Observation 16112627-729f-42ed-bf11-fb52fdf6bb03 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

TPTT: Transforming Pretrained Transformers into Titans Measuring Massive Multitask Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:46.194931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:46.194931Z digest=sha256:2d15038c8fb38f85adb1a0249ac2756f2bdb4ec603b170753a63cfbeeae14711

Observation d73a785d-e3d5-41d3-acdf-c13d4e76312a · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022.

TPTT: Transforming Pretrained Transformers into Titans Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:46.355340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:46.355340Z digest=sha256:8ef7bb088d6b42da15680aa8cd601a561e5f912d935191735817dd3f9879169e

Observation 0c76ca80-4926-4925-a5b0-0df37af3383e · outbound

This paper cites Jiang and all Mistral team.

TPTT: Transforming Pretrained Transformers into Titans Jiang and all Mistral team

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:34:50.786556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T23:34:46.464066Z digest=sha256:3f46e48d4e20550c8eb7e8c95a63d70fcd48d3fbfb521a94abdf97a6e3f0d491

Observation a9882323-0095-4026-949c-8c7e9223a4a9 · outbound

This paper cites Transformers are rnns: Fast autoregressive transformers with linear attention.

TPTT: Transforming Pretrained Transformers into Titans Transformers are rnns: Fast autoregressive transformers with linear attention

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:34:50.590682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T23:34:46.663137Z digest=sha256:e71863ddaf1891c73aec86da7a5f5c6d2d505eb7f9188d4dc9556bc2f9ee4b2f

Observation cc375e24-1f4e-4dbb-818b-8f96341d2f05 · outbound

This paper cites Liger: Linearizing Large Language Models to Gated Recurrent Structures.

TPTT: Transforming Pretrained Transformers into Titans Liger: Linearizing Large Language Models to Gated Recurrent Structures

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:46.813241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:46.813241Z digest=sha256:1ccca5fae54e2925045f207e92b998c37b1b5c47c663037a78bdf026e7a70205

Observation 3e1bf545-039e-4d5c-a465-f7183bd2a320 · outbound

This paper cites Language Models are Few-Shot Learners.

TPTT: Transforming Pretrained Transformers into Titans Language Models are Few-Shot Learners

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:46.956734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:46.956734Z digest=sha256:32ef31ab20e4215ed9898b8f957b1b7b127de8dc1efaaf353cb46506827f7b47

Observation 72fc9516-f739-4581-990c-54f71b667efe · outbound

This paper cites Openelm: An efficient language model family with open-source training and inference framework.arXiv e-prints, pages arXiv–2404, 2024.

TPTT: Transforming Pretrained Transformers into Titans Openelm: An efficient language model family with open-source training and inference framework.arXiv e-prints, pages arXiv–2404, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:34:50.424664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T23:34:47.085056Z digest=sha256:2ec1061ca2711fcc345196281602a7acfde7483fc4ddb53519ffdae844e29a7f

Observation 40e04140-7b47-419f-b80c-d557b484fde3 · outbound

This paper cites Linearizing Large Language Models.

TPTT: Transforming Pretrained Transformers into Titans Linearizing Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:47.210623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:47.210623Z digest=sha256:880b2d7575cccea3a6e3af4612bc59b08ab94df360b12057084633dcac9238ba

Observation 0e10e554-38d7-4b63-9a61-215e768e025a · outbound

This paper cites The Illusion of State in State-Space Models.

TPTT: Transforming Pretrained Transformers into Titans The Illusion of State in State-Space Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:47.274822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:47.274822Z digest=sha256:f6e402f3cd76c835d9a05df2fd7ba87874f933e2c2a7615efb746be804ef27f3

Observation baa28aaa-37c5-4746-a0ed-4a4ef2921bcb · outbound

This paper cites OLMoE: Open Mixture-of-Experts Language Models.

TPTT: Transforming Pretrained Transformers into Titans OLMoE: Open Mixture-of-Experts Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:47.387361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:47.387361Z digest=sha256:855527240ef519b7f50af16b997d7e308df29154fcf714ed1ef4eaa0ef3cc675

Observation 100ac0ea-d60b-4b96-9712-1da2ac80d36c · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

TPTT: Transforming Pretrained Transformers into Titans RWKV: Reinventing RNNs for the Transformer Era

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:47.534750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:47.534750Z digest=sha256:17827cb8b87adef6d699e2e971469d896e1b7c92d6ec8ad05a2f8d3b421d5899

Observation 33f593a6-e103-4ee4-9028-e23b41f95325 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

TPTT: Transforming Pretrained Transformers into Titans DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:47.659435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:47.659435Z digest=sha256:19416c88bf5fd4934bebefb417aa54380a0da0dd3acb4e955b012acb82d75192

Observation b1170e07-3796-401a-8362-58fd1fed7b57 · outbound

This paper cites Deltaproduct: Improving state-tracking in linear rnns via householder products.

TPTT: Transforming Pretrained Transformers into Titans Deltaproduct: Improving state-tracking in linear rnns via householder products

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:47.823694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:47.823694Z digest=sha256:ae260e53f5f4111bbea632711da0e772a7e62aa050b2611e14134fca9654ca18

Observation d48bbe81-7b06-4aec-9d20-71c12f50bf49 · outbound

This paper cites Alpaca: A strong, replicable instruction-following model.Stanford Center for Research on Foundation Models.

TPTT: Transforming Pretrained Transformers into Titans Alpaca: A strong, replicable instruction-following model.Stanford Center for Research on Foundation Models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:34:50.187128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T23:34:47.897344Z digest=sha256:4058d228c98dc8683106db63af81bd92575fec01fac9c58fba033b68cbc74e0c

Observation dd9da59a-d08b-40c6-8004-d8d7967abfa8 · outbound

This paper cites Gemma 3 Technical Report.

TPTT: Transforming Pretrained Transformers into Titans Gemma 3 Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.045442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.045442Z digest=sha256:ff220a1e0a2a3495fad2683eb519edac592983be623401db5ef2dbadca68094b

Observation da556d34-e987-43df-9601-931953106965 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

TPTT: Transforming Pretrained Transformers into Titans LLaMA: Open and Efficient Foundation Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.149039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.149039Z digest=sha256:8bf72cf0e60486b7039fba6bd9b558e2f447a6c870eb5dac0cd10c23eb302d18

Observation 8a99ca83-3241-4da8-8e03-ded4c01afcb9 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems , 30, 2017.

TPTT: Transforming Pretrained Transformers into Titans Attention is all you need.Advances in neural information processing systems , 30, 2017

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.337710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.337710Z digest=sha256:5cf6b35b78c6e8e83cd598df9482db94a4ff3ab77549d8b55804220ae90cf6a0

Observation 74404f03-18cf-4dd9-bf45-ed16d4537a4e · outbound

This paper cites Linformer: Self-Attention with Linear Complexity.

TPTT: Transforming Pretrained Transformers into Titans Linformer: Self-Attention with Linear Complexity

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.427452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.427452Z digest=sha256:9a75df635a9089f80fee4e1719cd7c82fe8263139b0b8d849488a81dc4097205

Observation 1e1eccb2-dc13-48c8-b255-08e6364afbe2 · outbound

This paper cites Parallelizing Linear Transformers with the Delta Rule over Sequence Length.

TPTT: Transforming Pretrained Transformers into Titans Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.547568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.547568Z digest=sha256:f679c11cc4083b84fe018326cbd0c481e72611661dfe61ed5570fb9c96abb6fb

Observation df9f8a2d-516a-4340-833a-e8c274bb271a · outbound

This paper cites Dream 7B: Diffusion Large Language Models.

TPTT: Transforming Pretrained Transformers into Titans Dream 7B: Diffusion Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.663768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.663768Z digest=sha256:5ed010c55a75c57409218ad6b92fbfc6884c648436a8588c03bbb21c55c5471e

Observation 96bf899d-d7ca-43ea-b338-d469d97cd370 · outbound

This paper cites LoLCATs: On Low-Rank Linearizing of Large Language Models.

TPTT: Transforming Pretrained Transformers into Titans LoLCATs: On Low-Rank Linearizing of Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.726772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.726772Z digest=sha256:7e92b4ecc0e5cdfd5ae9f5f0d0c440bf416c88bc283a31c0458018767a80906a

Pith citing papers

Observation a810ea4f-6889-4f42-ae02-17d8861ab7a7 · inbound

Federated Nested Learning: Collaborative Training of Self-Referential Memories for Test-Time Adaptation cites this paper.

Federated Nested Learning: Collaborative Training of Self-Referential Memories for Test-Time Adaptation TPTT: Transforming Pretrained Transformers into Titans

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:43:51.239271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-20T23:41:14.887651Z digest=sha256:663f88c1d9f12c4e85ebee833c8817a43d69b01cc0a6bf2967a108c72fdd20c7

Observation 47d1460d-280d-4d5f-b7a2-e32be4c589f4 · inbound

Black-Mamba: Biologically-Inspired Leaky Accumulation for Conceptual Knowledge under Distribution Drift cites this paper.

Black-Mamba: Biologically-Inspired Leaky Accumulation for Conceptual Knowledge under Distribution Drift TPTT: Transforming Pretrained Transformers into Titans

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T14:02:35.245972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:02:35.245972Z digest=sha256:6e270cd8b5a775420e40f4d938222a27655b445ddf438111494da4d60a44e75c