Pith. sign in

Paper Citation Record · LEDGER

Transformers Learn Shortcuts to Automata

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 46 inbound Pith citation observations for arXiv:2210.10749.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2210.10749 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 46 of 46 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:02:09.520474Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

11
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e5b27847-6b6e-4326-974b-e19842184d96 · inbound

The Computational Limits of State-Space Models and Mamba via the Lens of Circuit Complexity cites this paper.

The Computational Limits of State-Space Models and Mamba via the Lens of Circuit Complexity Transformers Learn Shortcuts to Automata

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:09.520474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:02:09.520474Z digest=sha256:b37394cf471a6dae0f8316610303e0bbc12b975b7bb6f27e34a2f6121cd90c8a

Observation 60fa9eec-f448-4988-9be7-5d3a5ca47b61 · inbound

Neural Scaling Laws Rooted in the Data Distribution cites this paper.

Neural Scaling Laws Rooted in the Data Distribution Transformers Learn Shortcuts to Automata

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T18:30:15.551782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:30:15.551782Z digest=sha256:b5d74ec827648ee27a8f9b5c041910b59a6f1229936723bc357e5c5b74e363ef

Observation 6d58a7dd-489c-4d65-9e50-19c8cf1607bf · inbound

ICLR: In-Context Learning of Representations cites this paper.

ICLR: In-Context Learning of Representations Transformers Learn Shortcuts to Automata

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T23:24:01.543001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:24:01.543001Z digest=sha256:43f5827203f408812fedd981630c4c5e76e9965392678cb4f5100e07693c7639

Observation 823d6ef6-561a-42a9-9334-0ecadd2d20aa · inbound

Rethinking Addressing in Language Models via Contexualized Equivariant Positional Encoding cites this paper.

Rethinking Addressing in Language Models via Contexualized Equivariant Positional Encoding Transformers Learn Shortcuts to Automata

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-10T22:49:44.902337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:49:44.902337Z digest=sha256:4f4e1aba7578ce17969cf105b0cd165e2537b5a37b024a793510a567106019ea

Observation f4c809e4-c01c-4e6a-b896-a020c5cd77aa · inbound

Learning Spectral Methods by Transformers cites this paper.

Learning Spectral Methods by Transformers Transformers Learn Shortcuts to Automata

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:51.255890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:51.255890Z digest=sha256:77758905a010d111e68a5fc64846dde2610bdf9fa1bcad55f72e44a92d4ae046

Observation ca9ab04d-7cf6-481f-9477-c0ffab8f5b9b · inbound

Circuit Complexity Bounds for Visual Autoregressive Model cites this paper.

Circuit Complexity Bounds for Visual Autoregressive Model Transformers Learn Shortcuts to Automata

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:42:46.274584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:42:46.274584Z digest=sha256:4a0946af531891355704a67ae538399437a1dfd35c2f17cd474aa7be5366e208

Observation 276bfca0-e664-4254-95a6-d80ad59018ca · inbound

An Analysis for Reasoning Bias of Language Models with Small Initialization cites this paper.

An Analysis for Reasoning Bias of Language Models with Small Initialization Transformers Learn Shortcuts to Automata

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T05:28:51.580568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:28:51.580568Z digest=sha256:598700a10faddc9e5a9eded838c762ac0207bb1718355e4348650768415d74e2

Observation 700c731c-2051-4cff-83c0-ff51694388c2 · inbound

Transformers versus the EM Algorithm in Multi-class Clustering cites this paper.

Transformers versus the EM Algorithm in Multi-class Clustering Transformers Learn Shortcuts to Automata

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T17:13:25.317497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:13:25.317497Z digest=sha256:c9c309e90bff26b30acf35f9b8be5172daf269b63c1b7c711534b288cd524552

Observation e5ba0a7c-78ce-45a7-9480-151d78a5ac5a · inbound

Too Long, Didn't Model: Decomposing LLM Long-Context Understanding With Novels cites this paper.

Too Long, Didn't Model: Decomposing LLM Long-Context Understanding With Novels Transformers Learn Shortcuts to Automata

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:30:46.668688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:30:46.668688Z digest=sha256:ae7cd983e8f9a23ffd981a9f786fd9449655e992bc3bf621282664c71aa24a4e

Observation 5bae81ac-d125-41e9-bf4f-39ce966d6004 · inbound

Transformers as Multi-task Learners: Decoupling Features in Hidden Markov Models cites this paper.

Transformers as Multi-task Learners: Decoupling Features in Hidden Markov Models Transformers Learn Shortcuts to Automata

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:57.702995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:40:57.702995Z digest=sha256:6aec3ef48a6a22140e25673e5275c2502ab26d85a8726bcc8f3963ad2edf7ef5

Observation 48788e9f-4700-4a99-b50f-d5541c2ea74e · inbound

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning cites this paper.

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning Transformers Learn Shortcuts to Automata

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:40.932026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:40.932026Z digest=sha256:671519ce5dac44519497a3b062b1a5a4bd0d3136e1bfa7e9826d7ae82a09ba45

Observation 5dcb5bf9-8a9d-41b7-b476-5a3842da069c · inbound

Transformers Meet In-Context Learning: A Universal Approximation Theory cites this paper.

Transformers Meet In-Context Learning: A Universal Approximation Theory Transformers Learn Shortcuts to Automata

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:33:40.540930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:33:40.540930Z digest=sha256:1bfb0756c1e9429b87a4c36c8675a193938f4e801aecdda7e40c5b772d2f294a

Observation 57831585-3380-467f-adc1-c3cd17489a82 · inbound

Sample Complexity and Representation Ability of Test-time Scaling Paradigms cites this paper.

Sample Complexity and Representation Ability of Test-time Scaling Paradigms Transformers Learn Shortcuts to Automata

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:36.314306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:36.314306Z digest=sha256:d7aa88b4c71dcfc06eee0160a4f53241b2e71a72ec799b223d37a22c9a725b1a

Observation 4de1ab7e-9de2-452b-9150-05f154133b75 · inbound

A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization cites this paper.

A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization Transformers Learn Shortcuts to Automata

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T06:10:19.422720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:10:19.422720Z digest=sha256:c18f38d617416d2f981ae4d531bbd4e39291e29c048bdfc220520dff87ec62c8

Observation 9a3b1e9c-f6fe-44f1-9d6b-2972aaa0ce8d · inbound

Position: A Theory of Deep Learning Must Include Compositional Sparsity cites this paper.

Position: A Theory of Deep Learning Must Include Compositional Sparsity Transformers Learn Shortcuts to Automata

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T20:35:45.748309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:35:45.748309Z digest=sha256:f7c541ee2405de94413ee518661c0edb8ec1e053e0e4789b21bdd301eaca0c92

Observation 442c24ac-807f-4b9b-ad3c-065a2c9fe61d · inbound

The Serial Scaling Hypothesis cites this paper.

The Serial Scaling Hypothesis Transformers Learn Shortcuts to Automata

Reference 63

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T04:12:02.449702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-19T04:08:11.344622Z digest=sha256:6310d166f0b0c90a151fca61b7529bccbd778b0e0be0866c11235a3191012d7b

Observation 3f495db6-6e73-4c31-8b3b-5c85d0f97a6c · inbound

Rethinking Memorization Measures and their Implications in Large Language Models cites this paper.

Rethinking Memorization Measures and their Implications in Large Language Models Transformers Learn Shortcuts to Automata

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:54:32.921274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:54:32.921274Z digest=sha256:bd6e100fd7af44a53f27c281aa1e345558a656e53ee4122f36f8482795c62244

Observation 7fe3a340-0f92-4b32-8efc-207640ecebe6 · inbound

Scaling Latent Reasoning via Looped Language Models cites this paper.

Scaling Latent Reasoning via Looped Language Models Transformers Learn Shortcuts to Automata

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T07:43:11.757752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T07:43:11.620446Z digest=sha256:84ee7f705b71d83d7927939d3d2fad8ed9aeaaa04137880ec5bac169abdff983

Observation 6f347182-a335-4f9b-b658-b4214a5eb266 · inbound

Scaling Latent Reasoning via Looped Language Models cites this paper.

Scaling Latent Reasoning via Looped Language Models Transformers Learn Shortcuts to Automata

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-04T07:31:49.264041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:31:49.264041Z digest=sha256:d1823545d4ca460149ea3e65833512f9bd763c0d5068188e3dc72ddcb2842529

Observation 4f7ecfea-a4f9-41e1-a60d-c1d7d35c80a7 · inbound

Context-Free Recognition with Transformers cites this paper.

Context-Free Recognition with Transformers Transformers Learn Shortcuts to Automata

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T12:57:45.253986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:57:45.253986Z digest=sha256:1580a898077cc76e08c4bf52b37353c14e2dc47759deab37c3e9e404e836076f

Observation 9099aeb3-cb78-488f-8bc0-4eb1cb01e8bb · inbound

Learning State-Tracking from Code Using Linear RNNs cites this paper.

Learning State-Tracking from Code Using Linear RNNs Transformers Learn Shortcuts to Automata

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:46:42.371401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T21:46:14.287966Z digest=sha256:a271403acf13600aa6dfc19ebf07c2e3a8cdd894f05f5a90d1c6a6308c59cf89

Observation 58a2416f-22c8-488e-a97f-6b0a15e7ff13 · inbound

Learning State-Tracking from Code Using Linear RNNs cites this paper.

Learning State-Tracking from Code Using Linear RNNs Transformers Learn Shortcuts to Automata

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T23:07:48.566812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:07:48.566812Z digest=sha256:f2285962e305878a0e50ecd2b9e91a59b9ac452c4de14c61bdf1f247911d12f4

Observation e1d9394b-eafa-4d0e-98f4-fd890b5ea548 · inbound

On the Emergence of Implicit Curriculum in RLVR Learning Dynamics cites this paper.

On the Emergence of Implicit Curriculum in RLVR Learning Dynamics Transformers Learn Shortcuts to Automata

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T23:11:54.933099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T23:11:54.933099Z digest=sha256:08d2276465940a28fd332d33581f41195f97a9deda207bc80b0ba690ebc1349d

Observation 3e08dc52-cc42-4d32-85cb-c2d2bb4baaef · inbound

The Recurrent Transformer: Greater Effective Depth and Efficient Decoding cites this paper.

The Recurrent Transformer: Greater Effective Depth and Efficient Decoding Transformers Learn Shortcuts to Automata

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:21:04.612914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-09T22:03:35.260373Z digest=sha256:01d4dbcac40b0ee6ce1fb39b21e0a5e8a10b58a9fc6715852ebede4544ecc340

Observation e081b988-246e-432b-95e6-ca9090bba7eb · inbound

There Will Be a Scientific Theory of Deep Learning cites this paper.

There Will Be a Scientific Theory of Deep Learning Transformers Learn Shortcuts to Automata

Reference 243

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:21:09.052414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-09T20:11:17.616190Z digest=sha256:a82fe679c1b7bc16d592554c414e33c3cd1e7ab3446f9e6ee4011c93333347ae

Observation 1ff35386-6736-469d-9aee-64d91b4d0c21 · inbound

Do Neural Operators Forget Geometry? The Forgetting Hypothesis in Deep Operator Learning cites this paper.

Do Neural Operators Forget Geometry? The Forgetting Hypothesis in Deep Operator Learning Transformers Learn Shortcuts to Automata

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:41:10.875342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T14:43:03.403254Z digest=sha256:d588950dca43abba4d35d0d95b0c3803395d128bee2e20e988c418ffa8e22809

Observation 81112096-4b4e-429a-8627-12c4a469f16f · inbound

The two clocks and the innovation window: When and how generative models learn rules cites this paper.

The two clocks and the innovation window: When and how generative models learn rules Transformers Learn Shortcuts to Automata

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:18.663325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-12T03:15:45.257213Z digest=sha256:fc6e0d2eb4e29054fa052c232230c0089fc1af1a1d0f5c9bfffb05ec302087ce

Observation 2b1f22dc-b60d-4288-a0f4-7c625c13d22f · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Transformers Learn Shortcuts to Automata

Reference 89

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:17:54.590712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:f9bf719548cd40a2764562fc9cc7ca7250432750f2b6ac805545144b7e05b38a

Observation a05b1fdf-2fee-4367-899a-225eb7f54f39 · inbound

A Sharper Picture of Generalization in Transformers cites this paper.

A Sharper Picture of Generalization in Transformers Transformers Learn Shortcuts to Automata

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:03:59.311962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T06:01:03.524290Z digest=sha256:857938eba04fac54e7b2d1f4a81a39237d20561e6a22b4c07792db8986373418

Observation 9ed966bd-98b2-40e6-af23-d6c7b6e06202 · inbound

A Sharper Picture of Generalization in Transformers cites this paper.

A Sharper Picture of Generalization in Transformers Transformers Learn Shortcuts to Automata

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:24:57.098974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T17:22:30.861739Z digest=sha256:f7b4d1239d85a1577ecdb3a65a7cc0f86564e295a8e26ac7a033c248f76b29ad

Observation 40218d6f-a608-4412-af61-ff896c6b4655 · inbound

Lost in Tokenization: Fundamental Trade-offs in Graph Tokenization for Transformers cites this paper.

Lost in Tokenization: Fundamental Trade-offs in Graph Tokenization for Transformers Transformers Learn Shortcuts to Automata

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:04:41.429666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-22T07:03:52.438466Z digest=sha256:249966c58351f42e013f0c2a33a016c80cd7c8279dc3da3cd91fe6a477b235cd

Observation 12ba1f3d-d77c-4cfc-917e-08c45af40ffa · inbound

Do Language Models Need Sleep? Offline Recurrence for Improved Online Inference cites this paper.

Do Language Models Need Sleep? Offline Recurrence for Improved Online Inference Transformers Learn Shortcuts to Automata

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:43:59.505955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T21:37:51.638904Z digest=sha256:702d0a40eef706fd0038c9d43d2ce4a56a045040113bc759e4ded60ee47ca9e2

Observation fd42cb4b-53fa-46c5-b0b6-a29b9b8c7dfa · inbound

Transformers Provably Learn to Internalize Chain-of-Thought cites this paper.

Transformers Provably Learn to Internalize Chain-of-Thought Transformers Learn Shortcuts to Automata

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T14:33:30.623410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-29T14:29:10.010212Z digest=sha256:1fbac4bfe30602f9c2160759dfb21d8dd6c8246376e85fb087008433bfa07b8e

Observation 49fef9f7-1329-402c-8c38-1bb5fc5ac7c3 · inbound

Agentic Transformers Provably Learn to Search via Reinforcement Learning cites this paper.

Agentic Transformers Provably Learn to Search via Reinforcement Learning Transformers Learn Shortcuts to Automata

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:42:49.952878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T23:26:28.158991Z digest=sha256:a4e915b9664bfb05a072ba50628d8ee5ffd50c2263252c74c9de7c8ac8fd88df

Observation 6647f38c-8077-4a2f-b088-2ae9909aa345 · inbound

A Close Look At World Model Recovery In Supervised Fine-Tuned LLM Planners cites this paper.

A Close Look At World Model Recovery In Supervised Fine-Tuned LLM Planners Transformers Learn Shortcuts to Automata

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:46:26.378690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T11:36:14.896698Z digest=sha256:e9b8c3796827a12f1ece76d12479b724a92a5de8eecd210aa217f6167313f1bc

Observation 56160b2b-987f-42d5-9049-b5788982be3d · inbound

Pretraining Recurrent Networks without Recurrence cites this paper.

Pretraining Recurrent Networks without Recurrence Transformers Learn Shortcuts to Automata

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:26:56.531884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T02:09:01.018909Z digest=sha256:0005ef8fa441cf99ed597e0bdd3427b2f81981f3bfc54deebf3643893e54e091

Observation a00123fd-aaf9-4bc1-b3c7-0d7721d119c9 · inbound

Pretraining Recurrent Networks without Recurrence cites this paper.

Pretraining Recurrent Networks without Recurrence Transformers Learn Shortcuts to Automata

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-02T12:20:55.718517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:20:55.718517Z digest=sha256:58c8fbe1027e9d770d494c116ba1d67066713cbff9f1de52baded243829e9621

Observation f2a342aa-1bca-4cae-9ab3-ec3e11edd33e · inbound

A Systematic Study of Behavioral Cloning for Scientific Data Annotation cites this paper.

A Systematic Study of Behavioral Cloning for Scientific Data Annotation Transformers Learn Shortcuts to Automata

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-06-29T16:23:39.049076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-29T16:23:08.402194Z digest=sha256:41bac0e54921af46872e174afe2ffe2a92a217d40eb5d4485cbbbb07dbcf13da

Observation 79f9775e-4088-4922-a345-9a23f911205b · inbound

Learning Dynamics of Chain-of-Thought State Tracking in a Solvable Transformer Model cites this paper.

Learning Dynamics of Chain-of-Thought State Tracking in a Solvable Transformer Model Transformers Learn Shortcuts to Automata

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:39:05.223791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T21:49:03.896740Z digest=sha256:564995e65a7363148bba48221baf780dbfe0971280e28cb06ee26442f4044d0b

Observation 657ec5a5-be14-45a7-be74-8e2696e06f9a · inbound

Critical Percolation as a Synthetic Data Model for Interpretability cites this paper.

Critical Percolation as a Synthetic Data Model for Interpretability Transformers Learn Shortcuts to Automata

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:49:29.662839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T17:41:29.317167Z digest=sha256:e1078576752f029dcef2ce62633c0611d4bf6fd3559294f40f07e71dff3a4d23

Observation c4cc34ed-2a9c-4f98-94d3-6f293ebb5b2b · inbound

Structure Before Collapse: Transient semantic geometry in next-token prediction cites this paper.

Structure Before Collapse: Transient semantic geometry in next-token prediction Transformers Learn Shortcuts to Automata

Reference 147

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.073513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-26T05:14:07.208255Z digest=sha256:2a25c87e0e11b23971c543579995fab1dd4117a5b520f7d62467e7cb43622e01

Observation fd43a091-3977-4800-8a3e-e37d8cc561dd · inbound

A First-Principles Theory of Slow Thinking and Active Perception cites this paper.

A First-Principles Theory of Slow Thinking and Active Perception Transformers Learn Shortcuts to Automata

Reference 106

Resolution
verified exact
local_arxiv, observed 2026-07-10T11:37:03.236616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-10T11:32:24.374377Z digest=sha256:42e2ce1007bc73ea52e2a8b4f5e9a50a944c400facd81b6f48e05449fe4292fa

Observation c567fd28-cc10-46ca-9f8f-540a197b882b · inbound

When Does Reward Teach State? A Hidden-Automaton Instrument and a Group-Language Warning Signal cites this paper.

When Does Reward Teach State? A Hidden-Automaton Instrument and a Group-Language Warning Signal Transformers Learn Shortcuts to Automata

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T07:20:39.305688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T07:20:39.305688Z digest=sha256:f29af25227386b9eba854cb395ddf9f2ec59888e3ef8722d15039e324376903a

Observation 6460473a-1666-4202-bf6e-93dfefaca180 · inbound

Hierarchical Domain Generalization cites this paper.

Hierarchical Domain Generalization Transformers Learn Shortcuts to Automata

Reference 172

Resolution
unresolved
no resolver link, observed 2026-08-01T20:54:18.389511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T20:54:18.389511Z digest=sha256:0ede1f39151b3d5dc2a412cd25dac4d6c9519c1af05d51ce3e9468ff21d6b51a

Observation 643edcf8-3e44-44a7-946f-a011eac12d90 · inbound

Naju: A Native Discrete State-Space Model with Independent Retention and Writing for Long-Sequence Memory cites this paper.

Naju: A Native Discrete State-Space Model with Independent Retention and Writing for Long-Sequence Memory Transformers Learn Shortcuts to Automata

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T08:50:28.180150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:50:28.180150Z digest=sha256:8106932fc3f98d6788f50423c364d2bbd49602630485c1b584e76b62f7a7dc95

Observation caefcef9-4e90-4bb7-942e-547adae906e2 · inbound

Attention-based representations for multi-task computation cites this paper.

Attention-based representations for multi-task computation Transformers Learn Shortcuts to Automata

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T00:18:21.309970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:18:21.309970Z digest=sha256:2602d6c0427500f5c960cc07d98eb119845b6900d51704255e27d3433c670a28