Pith. sign in

Paper Citation Record · LEDGER

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks

As of 8 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2607.07884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07884 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T16:03:20.699893Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact4
  • verified fuzzy49
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a687248-9d56-477f-9000-e36bf26ed1e4 · outbound

This paper cites Neural networks and principal component analysis: Learning from examples without local minima.Neural Networks, 2(1):53–58.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Neural networks and principal component analysis: Learning from examples without local minima.Neural Networks, 2(1):53–58

Reference 1

Resolution
verified exact
doi, observed 2026-07-10T16:07:20.252525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:f3efed7e39553aa0df55b9a560aac7b01a105e6f1c832f17b46fe9939c6b2d69

Observation cd8e9f22-3e4a-41a6-b4d8-2973b66b3033 · outbound

This paper cites Gen , volume=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Gen , volume=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.562939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:b1a4a8240e47f9892d57d0f12e132bc693e7fe7945a03f01fa5d689f1d9b9f61

Observation da70ab83-7e20-47ce-87a7-ba88138088a3 · outbound

This paper cites Proceedings of the Thirteenth International Conference on Artificial Intelligence and Statistics , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the Thirteenth International Conference on Artificial Intelligence and Statistics , pages =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.565366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:4dc328060822dc87e2832043cfe70ca1ac966b28e314d3a6710e062e818dead2

Observation 4cabddbe-cf2d-404f-8b71-030109a7906b · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks International Conference on Learning Representations , year=

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.648305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:500635ed1e15e3a199f8e92014800d2277fc5e79e5d10e69af9b8a0519dfc623

Observation a55e118c-4e7c-489e-9e67-ca0ce54ca940 · outbound

This paper cites Exponential expressivity in deep neural networks through transient chaos , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Exponential expressivity in deep neural networks through transient chaos , url =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.633775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:fcefa7ee367ac5dcefbd1dfc5205b6a03b9f9d1f68f7d71739379c84a1ad4267

Observation 3023b154-5238-488e-9399-925e4cb6a0c7 · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks International Conference on Learning Representations , year=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.574459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:f6f9820e1a52238ea6664693d1cb88ccd62f2ef723b66ee3df7ba1849f5d19f4

Observation 930da198-b5d1-4d9a-ba3d-1e1d45a8b20b · outbound

This paper cites Zico Kolter , booktitle =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Zico Kolter , booktitle =

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.615801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:1c4ff3cd3b316d63c4357b0da6928a484c2404ef24dc518dd42ff2df74cff932

Observation f1a9f19c-89ea-47ae-a274-c3d03a3be811 · outbound

This paper cites Stable architectures for deep neural networks.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Stable architectures for deep neural networks

Reference 8

Resolution
metadata mismatch
doi, observed 2026-07-10T16:07:20.254811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:ceec4b193baed37b0d69f6afe8d8a62a3ed285eddce16e299fcfbc11a112573c

Observation d701b219-cb76-4ed4-b3e3-de792b432e17 · outbound

This paper cites an unresolved cited work.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-07-10T16:07:20.554030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:906133d6b5aca3e8d3115ba701264baeade8e63a31990755950e1e38983825a1

Observation 368a6c83-d58d-40e4-92e6-45324465ea90 · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks International Conference on Learning Representations , year=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.636631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:0ad2b5bd5b5bb6dc786e4f7aac56c243513ab6cfa100bf0777a4a1b5b45623e3

Observation 215594c1-c043-4858-a838-839edcf6f5e8 · outbound

This paper cites Dynamical Isometry and a Mean Field Theory of.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Dynamical Isometry and a Mean Field Theory of

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.576754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:ca8f8a6ebcf0541ba49721000512c90309cdd6ab073c0a15d65d7d98fbf89dde

Observation 687a7327-11a1-4804-8bfb-5b65337e3963 · outbound

This paper cites Zico and Koltun, Vladlen , booktitle =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Zico and Koltun, Vladlen , booktitle =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.578822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:852cbd9ba737bf7b4bddecbd99edb0841c686b743b4e20396268cf0564789443

Observation caf47eda-729a-4b56-a7bb-5bf20b8402a5 · outbound

This paper cites Mathematics , VOLUME =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Mathematics , VOLUME =

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.639508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:f0aa14273991b987c9f23d7031f44e6302d347c3de72a601a322013480534be5

Observation 5364c6ae-9596-4256-a392-cbd56ab974b0 · outbound

This paper cites Saxe and James L.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Saxe and James L

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.597830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:4a3a9a8498494276bb8590fd337f3f1f54eccf0159b3ebe79c2ead4d74ef6ee1

Observation 542b995e-ea7c-4f8a-9e8c-f410dc5b14bb · outbound

This paper cites Proceedings of the Thirty-Second Conference on Learning Theory , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the Thirty-Second Conference on Learning Theory , pages =

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.623306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:c42237aec76db4f974e02ccd19893bd2fda7e344cb1320c829a1c93454b1a0a1

Observation 5d1c52f6-8eaf-4568-a70e-bcd096aa43cb · outbound

This paper cites Implicit Regularization of Discrete Gradient Dynamics in Linear Neural Networks , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Implicit Regularization of Discrete Gradient Dynamics in Linear Neural Networks , url =

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.596450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:3c1cb5306c4b2ac045c9b8aebeba8cd1cdca913614bd276e54d3d04841e57bb0

Observation 62db237f-7d76-4f55-9e29-45aa6fe2f193 · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks International Conference on Learning Representations , year=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.635414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:2ed09cf72080376dae55b6ba326d49b6ce68722cd0852dccb8774274aa99d4bb

Observation 0f7e6eb4-8e16-4465-b038-60675834985f · outbound

This paper cites 2020 , eprint=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks 2020 , eprint=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.632623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:14e2a6704b538a0be7e47ecaa5a3116e26e0853b79452681222b7d872f3eab0e

Observation 2c768d4c-b83b-4417-a873-a3eb4d611d19 · outbound

This paper cites An empirical analysis of compute-optimal large language model training , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks An empirical analysis of compute-optimal large language model training , url =

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.651059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:937a4f0d5bd643e688baec898f874d3c4abc8c02e1214fe4665849254e445466

Observation 0808331a-0776-4776-b6c9-4c6bb6f3d52e · outbound

This paper cites Advani and Andrew M.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Advani and Andrew M

Reference 20

Resolution
verified exact
doi, observed 2026-07-10T16:07:20.258992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:9b4cfd844e5273a8c13b5e13b2bb93f8c808db2961d68faefc4026c5a6550371

Observation cdb7c0e7-f654-4cc6-9b6b-0e2940bf317a · outbound

This paper cites Proceedings of the 38th International Conference on Machine Learning , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the 38th International Conference on Machine Learning , pages =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.625378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:79e9be444713d924fb05dbdcbdd193c2fec5697f640510ebf073764ea3febc40

Observation 118a5d4a-ad92-4b95-bb86-2f29327de38e · outbound

This paper cites Exact learning dynamics of deep linear networks with prior knowledge , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Exact learning dynamics of deep linear networks with prior knowledge , url =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.617816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:c939f87bf630946a7b868716b028faf52d4db1bf5f2caace4025dd09afd34044

Observation c78c5037-e4ee-49f4-915d-11f55c8e7f59 · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks International Conference on Learning Representations , year=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.648494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:3a09e6cbc5e344d9806b93fd4cea2ebb13f431aaedd83af89598eb5d42a2cf4c

Observation bbd5238b-a552-47a0-aa00-e406fd7b7fa7 · outbound

This paper cites Proceedings of the 35th International Conference on Machine Learning , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the 35th International Conference on Machine Learning , pages =

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.563279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:2fe6217c6635f832e586a55c73c90f30e59dee8bfa6dbd1897efd064a4042d9d

Observation 1770b131-9be9-4037-8c7f-184194bf71b3 · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks International Conference on Learning Representations , year=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.567880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:822d75f33346036f28cb99e63ba016a8a674296f534f41ccfa8d7e80bfb41c8d

Observation 9f6c0a02-482d-4969-90d5-70a52e8dfbd8 · outbound

This paper cites Algorithmic Regularization in Learning Deep Homogeneous Models: Layers are Automatically Balanced , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Algorithmic Regularization in Learning Deep Homogeneous Models: Layers are Automatically Balanced , url =

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.645434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:bc93ca286624d4a7ce34069ad82a8f7a1e31448a8de1993b5147183a9c0c5dbb

Observation 630c31d4-2604-46c5-94c3-c55aaf8d83fc · outbound

This paper cites Neural Tangent Kernel: Convergence and Generalization in Neural Networks , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Neural Tangent Kernel: Convergence and Generalization in Neural Networks , url =

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.623055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:a1db77de8cb1e4fe5a23ef1844baafda78b3731026efe88d9e5b44f7a30c8bea

Observation 51dc76ee-d3e1-4251-866c-e713a86707a3 · outbound

This paper cites Learning dynamics of deep linear networks with multiple pathways , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Learning dynamics of deep linear networks with multiple pathways , url =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.630994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:5a8c0188e653616caae99a7115ee91d5ab799c52a219ef49df0a3fadf1df35d2

Observation f2cefb4a-00a7-4c9b-aeef-3e75caf8eaeb · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks International Conference on Learning Representations , year=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.625746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:cb51ba238cf7856933e002090ee7f2d84188e0076637014cd203b184974276f1

Observation dab66938-3f60-44a3-980b-ea98c1c28db6 · outbound

This paper cites On Lazy Training in Differentiable Programming , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks On Lazy Training in Differentiable Programming , url =

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.593037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:02fed6beb180ad3664d69411891b4c1cc3c122f4f62f12d9c5a5b1bc39fb7465

Observation 1c3d7d0b-8a43-4c0e-8dff-0344efb8a3d6 · outbound

This paper cites Proceedings of Thirty Third Conference on Learning Theory , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of Thirty Third Conference on Learning Theory , pages =

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.572433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:b60f26bf5999fac5acaddfb1d514385ffb5f6145a56f2e253a7cbead9b0dd41d

Observation 4c3c4fcd-d375-4153-9275-14ce1849f4c7 · outbound

This paper cites Proceedings of the 37th International Conference on Machine Learning , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the 37th International Conference on Machine Learning , pages =

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.588362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:225a90a9b259760a1f0444a49d8c647a864cea92c67ee0b9dd600d02b7bd0d12

Observation 30a82447-6799-47ae-b1a5-dddadb43daf8 · outbound

This paper cites Saddle-to-Saddle Dynamics in Deep Linear Networks: Small Initialization Training, Symmetry, and Sparsity.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Saddle-to-Saddle Dynamics in Deep Linear Networks: Small Initialization Training, Symmetry, and Sparsity

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:07:20.385724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:e1cbce15fabc05be61d8d6cb831fda0ea2396d707ba7d1367ee9f8595049fd81

Observation 4a64eb52-b03e-4b8a-9c94-e9858b371e5f · outbound

This paper cites The Shaped Transformer: Attention Models in the Infinite Depth-and-Width Limit , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks The Shaped Transformer: Attention Models in the Infinite Depth-and-Width Limit , url =

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.645229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:3b570e92e7f199a7299debabbf8afc902a192cde1af7087d1bcdb18d0125837d

Observation c8af1a34-01bf-4dca-b7fd-1f047e3c2486 · outbound

This paper cites Journal of Machine Learning Research , year =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Journal of Machine Learning Research , year =

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.627302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:c6c659e44ce04e22ab8a48aae98bb519e36dc3339fcafa082abce7078e71fd9f

Observation 80506e23-a505-4263-b3f2-3156c69913d3 · outbound

This paper cites The Eleventh International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks The Eleventh International Conference on Learning Representations , year=

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.601147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:4b3a8aca77a97975033f105f9c9f92e4e6d697c1b3845b2c3b0af38852438d38

Observation 24837d75-761d-4b7a-aa8e-adf4ebe0b17a · outbound

This paper cites Saddle-to-Saddle Dynamics in Diagonal Linear Networks , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Saddle-to-Saddle Dynamics in Diagonal Linear Networks , url =

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.599906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:ffd1299ac57723372bdd2cfa130515bb96eed312b3afd4c7769d57ed75f02b2e

Observation b2b9b646-b713-4f02-85f9-f1491fb2255c · outbound

This paper cites Proceedings of the 40th International Conference on Machine Learning , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the 40th International Conference on Machine Learning , pages =

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.618117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:49485954a92a4de5b55779324e21e19ac6ac8057c91760acb9dc04ff413f57af

Observation cb90fa70-3ad6-45dc-905a-4937f474621d · outbound

This paper cites Transactions on Machine Learning Research , issn=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Transactions on Machine Learning Research , issn=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.650648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:90241ddeb8dee302c6d6ab5145202d8a0b175af72468d078c230007e1347b794

Observation 04b7367e-3059-4e64-8e16-34d22b0b6b89 · outbound

This paper cites 2023 , eprint=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks 2023 , eprint=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.613455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:8e728df17983ddb07b66c4ad8511a8e3cd69b21a7cb24c5c1eb875450257ea14

Observation 44cbecbe-89e3-4a3e-a588-7c192c8a9c6f · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the 41st International Conference on Machine Learning , pages =

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.547835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:0c85aaa9d4258b9bd49ce0b7ed300f4f953cc233225f927d77c5a84560b0d52b

Observation 854a339b-c844-483d-8450-4c90eb41986a · outbound

This paper cites The Fourteenth International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks The Fourteenth International Conference on Learning Representations , year=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.545806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:d86469a0635a09a2c717295406544b3548dc17a40a2339cc49aace5871daf01d

Observation e5382b50-e944-481b-9f89-8e648a2537b3 · outbound

This paper cites Scaling ResNets in the Large-depth Regime , journal =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Scaling ResNets in the Large-depth Regime , journal =

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.611211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:c723414629ea3413efe2336c0a7ec5f79168cfc63f8f3e56ff0dd38b6bfc728b

Observation 8be7c1d1-ac8c-43a6-b313-1d83b08f127e · outbound

This paper cites doi: 10.52202/079017-1130.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks doi: 10.52202/079017-1130

Reference 44

Resolution
metadata mismatch
doi, observed 2026-07-10T16:07:20.257203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:d7c1ff4b860bc251ecb5e02116ceda23eb7db55131c6d53e8553bd6d9b682b4d

Observation bfada41c-7cc8-47ae-af52-c37df463f788 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks The Twelfth International Conference on Learning Representations , year=

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.629956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:f7461251f930b746a4c0e18375eced32875cc932e6fa439b48bd28087da6104c

Observation 0c2dcf6f-b129-4f30-bff3-be58607137a6 · outbound

This paper cites Super Consistency of Neural Network Landscapes and Learning Rate Transfer , url =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Super Consistency of Neural Network Landscapes and Learning Rate Transfer , url =

Reference 46

Resolution
verified exact
doi, observed 2026-07-10T16:07:20.249887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:bd37081313e5b447c05aa527f76aad54ad3b3a41b7884952fd353aa74d77e5bf

Observation b6636d3b-b8b5-4da3-92b4-879f77d677eb · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the 41st International Conference on Machine Learning , pages =

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.550520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:e595bace4f4734853b4aeb1f40d866329809f2cd113df267794b3c727aa18a3f

Observation bb609da8-58b9-4e95-b3c1-f6fab75a7a34 · outbound

This paper cites Tensor Programs.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Tensor Programs

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.654152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:62e138f934c7984bdfcb947e1aedf39ae7706ca49d5d2974a04ca7aab974ec42

Observation 964467f7-6524-4f2d-bd4b-506086066732 · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks The Thirteenth International Conference on Learning Representations , year=

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.642918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:31d3ca40453658dee88b7cd9094455f0d0f531a22573eef945deb74904ff5ff0

Observation 625120dd-4aa9-4ff5-bb31-0b5ddab44eba · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks The Thirteenth International Conference on Learning Representations , year=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.543545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:ee60a44ddf3cce73567abde8ac395db391dd5294c110856506d9f58a55eb6b9e

Observation bfeb656e-49dc-4958-ac4e-3d5d0ad6f6ec · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks The Thirteenth International Conference on Learning Representations , year=

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.591135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:664e120531578634f82ba17b559b431d42fc210f7054971d64926e0204152cff

Observation 86c5296c-e948-4b2d-8ea9-14071eb25c66 · outbound

This paper cites Proceedings of the 42nd International Conference on Machine Learning , pages =.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks Proceedings of the 42nd International Conference on Machine Learning , pages =

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.620548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:3fe4a5cd14928bc4bfb7b36f0d951dfba06db112e0f65982810d5cd3b2febced

Observation 1cff7f7d-b18d-4d5d-93ee-c4d68a0c0e0c · outbound

This paper cites The Thirty-ninth Annual Conference on Neural Information Processing Systems , year=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks The Thirty-ninth Annual Conference on Neural Information Processing Systems , year=

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.574296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:0321fcd6a7b03a1cec0160c52d3cf2cb23935f7a9a6a7296e7d8bf5ab72075c0

Observation e8cd5fa3-3c66-4e92-9a32-0a27643bdc73 · outbound

This paper cites 2025 , eprint=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks 2025 , eprint=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.579023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:a169b9a7ab852481c2eebc02528fb9fd2e1be286992ab0eca39569b16ad8b2c7

Observation d08d0956-90bd-4334-accb-62310e375adf · outbound

This paper cites 2025 , eprint=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks 2025 , eprint=

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.642388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:506bb60ac8dd1586c3781b551b3f8eb6366859b716b71fab41223aa1abf33ac0

Observation 9f486c21-df81-4f6f-8993-cd06ac6ae73b · outbound

This paper cites 2026 , eprint=.

Optimal Learning Rate Scaling Depends on Data in Deep Scalar Linear Networks 2026 , eprint=

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:07:20.640679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-10T16:03:20.699893Z digest=sha256:991ea1ce5188d56bf720624fe81d8f6410a5bf9919884837f7182d122fb7fb47

Pith citing papers

No inbound Pith citation observations are available.