Pith. sign in

Paper Citation Record · LEDGER

Optimal Training-Time Scaling in Gradual Adaptation

As of 16 August 2026, this Paper Citation Record lists 100 of 107 outbound references and 0 inbound Pith citation observations for arXiv:2608.04927.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04927 v2

Coverage vector

measured 100 of 107 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:23:55.291013Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 107 outbound references displayed

  • verified exact4
  • verified fuzzy55
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e23d139f-1991-478d-b669-65b08eb3a25c · outbound

This paper cites Neural Networks , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Neural Networks , volume=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.951480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.951480Z digest=sha256:fa8bf546e8c83ac435d38900f3ba877a7a6e28382094491ef5a537ba2d30659b

Observation 7a282466-fb66-4004-b1e4-f155f8454736 · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence , volume=.

Optimal Training-Time Scaling in Gradual Adaptation IEEE Transactions on Pattern Analysis and Machine Intelligence , volume=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.957471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.957471Z digest=sha256:3259def67e12309ab466d4c927d1049fb1449b9d063dd6d7447edc66b5f416cb

Observation 6d1d07d9-1105-4b79-ad3a-65696087eac6 · outbound

This paper cites Proceedings of the 37th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 37th International Conference on Machine Learning , series=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.961685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.961685Z digest=sha256:1936c5eb688fdfa4a2df9c77fa6831802baee9a5a98d7d513da37ec3e2a230cf

Observation 3fcc8628-8b6b-4a8c-a121-7c461c655c90 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.965989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.965989Z digest=sha256:da90ee2720432e9d90e9a226ee2a7b671014cedf302eed7142608bb3675d6ee0

Observation 6125fc55-c75c-435d-a75a-cffe541fb413 · outbound

This paper cites Proceedings of the 39th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 39th International Conference on Machine Learning , series=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.970044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.970044Z digest=sha256:f53b7c52484cb8089af5902689e99ca1d664825bac7e12e8ed0dc05e6cb061cc

Observation 50435ea4-65a9-41b5-a061-14af19d3ba2b · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.973946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.973946Z digest=sha256:85b12e684febb306ef43a4e84f0a96d44fa62a7327b2fa69ed8956d47b7dadc1

Observation 82b4073b-100c-475f-8d38-686a70e53b46 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.978059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.978059Z digest=sha256:ec4fb8fca0a9d6c7defbe59ecf925bfb19d2d446cd6dce882f025cdf2ede7b2a

Observation 71bc4d43-8295-4b69-818e-e173224bb37b · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.981486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.981486Z digest=sha256:23879f1bdf9ec018ecf34db78a10c51f3fc003ef7cd8cd00968f3ffcafcbc7b8

Observation 80d2337f-8934-4f0c-81c9-2f90be31c8b3 · outbound

This paper cites IEEE Robotics and Automation Letters , volume=.

Optimal Training-Time Scaling in Gradual Adaptation IEEE Robotics and Automation Letters , volume=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.984888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.984888Z digest=sha256:b1645da364b313ed4e940bfc8eb122426f226bf6072b151793ee8af048419a70

Observation e793334b-7eb1-4cb4-be75-763fa4db1626 · outbound

This paper cites 2025 , publisher=.

Optimal Training-Time Scaling in Gradual Adaptation 2025 , publisher=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.988603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.988603Z digest=sha256:876f0476449b5dafb9b92900e70d99bb5c81f7df527f6e80126bbd5c9e91d824

Observation 79410c6d-b280-4e80-b3c5-428bba7960af · outbound

This paper cites The Eleventh International Conference on Learning Representations , year=.

Optimal Training-Time Scaling in Gradual Adaptation The Eleventh International Conference on Learning Representations , year=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.992100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.992100Z digest=sha256:ac44ad1908efc16d389f0633782fd291c0718f27bc510836febf8b1fb8a4daf3

Observation 778a81ab-3bf9-49b0-b3a8-df6fdc3026a0 · outbound

This paper cites Proceedings of the 34th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 34th International Conference on Machine Learning , series=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.995740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.995740Z digest=sha256:fa3452e43bd5937055aa70b59f1a0df992f30e4aef94c8dbdf71f22127fde8a2

Observation 4a65cce3-68b0-40eb-b4ef-ad8fd3237142 · outbound

This paper cites Proceedings of the 39th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 39th International Conference on Machine Learning , series=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.999264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.999264Z digest=sha256:72018bd29d46951a97d256958eeb7a69ce02c5afb3eeb93939f1b7ee676dd868

Observation b33f205b-2587-45fe-b14f-f02654ca301d · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.002614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.002614Z digest=sha256:4761052d6e9564e70ffffa5bef7880db4cffb600b97dc825da3fa81a220a5697

Observation cc1cdec2-fe01-464d-9920-b033d85110a2 · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year=.

Optimal Training-Time Scaling in Gradual Adaptation The Thirteenth International Conference on Learning Representations , year=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.005983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.005983Z digest=sha256:6f3861044a52f7d693e027421cc8780af066a6b58fadf9981c57da3a13cb1a1a

Observation b4176eaf-f84f-407f-b9a4-9200126447ac · outbound

This paper cites Constructive Approximation , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Constructive Approximation , volume=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.009041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.009041Z digest=sha256:90222d7d80cd7f070d91203798312037dea44f48f6d054841e88bbb50dc6f5a8

Observation 28723244-f315-412c-8e3d-d17372ff3acf · outbound

This paper cites Journal of Machine Learning Research , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Journal of Machine Learning Research , volume=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.011834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.011834Z digest=sha256:acfebfe3cc72a01c757e9be5fa0ccf84f7927a08b1ba62453937e09223ae3ead

Observation bf8d2157-ba62-4526-88fe-550b97a6beda · outbound

This paper cites Proceedings of the 22nd International Conference on Artificial Intelligence and Statistics , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 22nd International Conference on Artificial Intelligence and Statistics , series=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.015260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.015260Z digest=sha256:10b6e00105f76f42126d2c7956c43e21063b9726c548f7b638b2646d0b043268

Observation 24150688-c0c6-4dda-a958-1f92ea3cc9d5 · outbound

This paper cites Proceedings of the 33rd International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 33rd International Conference on Machine Learning , series=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.018585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.018585Z digest=sha256:1f6bcfc1eafcbcfc3001a6328a758c5f24b41655f1497fbd3cb22274f24a3cba

Observation 8b943b9d-b7c7-4560-b07a-2f8dd14a64d3 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.021522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.021522Z digest=sha256:f1888a0174268238f336bb3df9ef26b314a5b66c35f407c530b76bc86d04e9f2

Observation 1ae1ee5b-0bd6-4b00-b173-6f0f6f9196d5 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 41st International Conference on Machine Learning , series=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.024692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.024692Z digest=sha256:82a571cb913e785cfa6433df5f87c90f33bb710428ee4c8af4870b649a8d0e7a

Observation 8cf76e94-e1cb-4a1e-bd21-856ad541e518 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.027910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.027910Z digest=sha256:156a14af254361ab1cf7ac6592390c8a0b79109892c7290ec7aff28a7bdaa6ca

Observation c11a9925-a61c-4d91-b5d6-575ac9889460 · outbound

This paper cites Proceedings of the 2nd Conference on Lifelong Learning Agents , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 2nd Conference on Lifelong Learning Agents , series=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.030952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.030952Z digest=sha256:6e616cc8e94e17afa4e6f5c20229be3c05abec4572026802954f8941b751bece

Observation 8e5d90de-de4b-4aaf-9776-b05eea297ec4 · outbound

This paper cites Proceedings of the 42nd International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 42nd International Conference on Machine Learning , series=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.033985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.033985Z digest=sha256:17650feb96e934ec9f44a65f667738ec7adaa3f916fff9d988152c21988aaff0

Observation 42fb60bd-eb1d-49df-88a4-e812f61aebba · outbound

This paper cites Proceedings of the 35th Conference on Learning Theory , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 35th Conference on Learning Theory , series=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.037524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.037524Z digest=sha256:9b78c4ef805b60966fbb8575c209670767ca9a65408066912db2077086d9d7ac

Observation 0bfc006c-c88d-4bd6-b4d3-9d91718a74f4 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.041086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.041086Z digest=sha256:15a7e3aeb79ee3b352e2984d93de236323a4427dc8a838c8b0655e65ac0dc06e

Observation fcc5815e-fcb0-40bc-a5de-1d755b6e7484 · outbound

This paper cites From Continual Learning to.

Optimal Training-Time Scaling in Gradual Adaptation From Continual Learning to

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.044205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.044205Z digest=sha256:5863652b0122d126629f355f9f2db252ba066bd1e926b0459cd72f1393b0ac57

Observation ee320a5c-54d4-42a2-9011-38725bb6a8ce · outbound

This paper cites Journal of the Physical Society of Japan , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Journal of the Physical Society of Japan , volume=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.476394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.047395Z digest=sha256:dff9a2555f5cb05a28f37467489afbc9584bd6e60fbd55271505efbc28f1165e

Observation c7381799-d744-44a3-b69a-239176c81b92 · outbound

This paper cites an unresolved cited work.

Optimal Training-Time Scaling in Gradual Adaptation Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:23:56.467980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.050567Z digest=sha256:934dae397a5cc8d11650bf7f18dd52244f14695edff6b3078a6d2bfaf0262f6d

Observation d27339df-c071-47d2-b9ba-f1b2336d30cd · outbound

This paper cites Communications in Mathematical Physics , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Communications in Mathematical Physics , volume=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.458708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.053770Z digest=sha256:041685cc5b6d681577597a1a55a43a13d0caa44fc52f6a86fa6006c66258471a

Observation c5b818af-c52f-40a9-9308-9a8e5ba2c282 · outbound

This paper cites Journal of Dynamics and Games , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Journal of Dynamics and Games , volume=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.449281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.061270Z digest=sha256:f44e20dd129a4c80a12efe9db483d6f10f5570523bebffcc7c66289174e5feae

Observation e0536149-9791-4654-909e-8c130c9f8f53 · outbound

This paper cites Proceedings of the 26th International Conference on Machine Learning , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 26th International Conference on Machine Learning , pages=

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.440004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.064990Z digest=sha256:05da46af893631f7f15e7d0dffbc9f179d694a90c2630496328ea896d0d00954

Observation 5e2c05b4-f0ce-4ee3-8c72-c5efb30900e5 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.429766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.068951Z digest=sha256:dcf084486dea343e4be68dce2f90c8a3619396267c46dfedf16c7c9932bbc7e7

Observation a3d88299-68a2-4e5e-acba-c6d9040b0ebc · outbound

This paper cites Proceedings of the 35th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 35th International Conference on Machine Learning , series=

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.419862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.072405Z digest=sha256:caa796ce2a941ea794bd09d736b58adb222e268c13452594a47ee517d345f4b0

Observation b48ecbe0-5187-4281-b5bf-d5962d64dfcf · outbound

This paper cites Proceedings of the 36th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 36th International Conference on Machine Learning , series=

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.409412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.075631Z digest=sha256:2e58a0383229f0fd4a028c43ceec1251a752256e8406e74aedc911f78b7be2d4

Observation 0bce7133-5bdf-4270-a477-480d3037b8f5 · outbound

This paper cites Proceedings of the 37th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 37th International Conference on Machine Learning , series=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.399323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.078849Z digest=sha256:30b8a48bd3e6d2efc1f0e2923b53b75072f607ec4cd55c63f3fd4b3178cd874e

Observation abf703a8-30b2-4a38-8f41-ddbc38c79f19 · outbound

This paper cites 2021 , url=.

Optimal Training-Time Scaling in Gradual Adaptation 2021 , url=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.388724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.081993Z digest=sha256:78a63ea28f76137391990df0bc212d019202a4bf875ea3a0b86e865a161a3dfc

Observation 22ed719c-ec9a-482e-8698-58d4def3ff01 · outbound

This paper cites Proceedings of the 39th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 39th International Conference on Machine Learning , series=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.378213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.085529Z digest=sha256:486f3a18ee5a151147b4ab3ccfcb4cbc343f239d17f6471a81fdd5ba04c1e2b3

Observation 4a862d50-9737-467d-b0a8-d46b6e8e50c6 · outbound

This paper cites an unresolved cited work.

Optimal Training-Time Scaling in Gradual Adaptation Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:23:56.367407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.088866Z digest=sha256:f8cf53c958e3c541a9e3f67f932a89b66dcb93aa0d874a2133d9e1528c3c65e4

Observation 12ba1344-cd11-4d56-bcb0-fa9718b68417 · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Training-Time Scaling in Gradual Adaptation International Conference on Learning Representations , year=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.357236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.092702Z digest=sha256:f4b009925fc3ab9798b5eade717a24de83435cc48ec7004796232257ff3f6f48

Observation e077afb1-fcca-4305-b605-444d669c55c0 · outbound

This paper cites Machine Learning , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Machine Learning , volume=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.346734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.096213Z digest=sha256:a2ce23c78e87705d3ae229465e5170dee99264f2bc73a2738a9b7cee8b4da891

Observation 52125504-a4da-4673-a341-6e68c74ada28 · outbound

This paper cites Journal of Machine Learning Research , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Journal of Machine Learning Research , volume=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.099569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.099569Z digest=sha256:b15d00d82e3a5a7f57c706fe964b73921e2d0bd05c21af44a7a8b87d0312bcae

Observation 579a2bfe-8a9b-4cb5-892d-b76d3bf39c1e · outbound

This paper cites and Darrell, Trevor , booktitle=.

Optimal Training-Time Scaling in Gradual Adaptation and Darrell, Trevor , booktitle=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.331512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.103037Z digest=sha256:64981e6c5445cf7b743fd4055a90759cabc5ba1f62ec9bed2508a41ceee1b11a

Observation 1697a3b2-e63e-4565-8b87-0c89df3b5285 · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.322673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.106482Z digest=sha256:da93f63c974f57e6c2a421cf1344a6f9e324c6f8006a8dcfee3b876d6f2e4300

Observation 7d3cb934-a885-410e-995c-dcdd03f2eb6c · outbound

This paper cites Proceedings of the National Academy of Sciences , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the National Academy of Sciences , volume=

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.313575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.110104Z digest=sha256:938e655f2e8587bdd24bd1a260d8b6803b8f49d528cfb20fb1bc2cc304e8405b

Observation ec8791c9-b6df-4e94-a37d-107286cefb90 · outbound

This paper cites Proceedings of the 34th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 34th International Conference on Machine Learning , series=

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.303863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.113652Z digest=sha256:d463cbf45085170687d23aec95c98c482b4e5e28e32c571e4ef5840b35f3de63

Observation dbf5d8cf-6174-4948-8eaa-3029c5e32ae8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.116812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.116812Z digest=sha256:22038cc4e1cb455eb5de00fdee6600c5ae5a6c16e5c54fe766f33c003263d6fa

Observation f885f70f-dd49-4e99-b0ce-30488489f77e · outbound

This paper cites Efficient Lifelong Learning with.

Optimal Training-Time Scaling in Gradual Adaptation Efficient Lifelong Learning with

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.120123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.120123Z digest=sha256:1685a8e01479e72e9ad94668e2153294275c299e527262f31299f96aee82c053

Observation 0385501a-c298-4012-a4f4-a28d5958e14f · outbound

This paper cites Proceedings of the 23rd International Conference on Artificial Intelligence and Statistics , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 23rd International Conference on Artificial Intelligence and Statistics , series=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.279286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.123655Z digest=sha256:e41d30d18c4bce6df23c5b610b36d11ffda026c17dbb57f86564060b577030bf

Observation aceb1571-d2a2-4446-a9d3-ec04bb57153d · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.268028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.127136Z digest=sha256:2de0a60270e96862a29ac43e4fbb4986efea51bf9e7e64eab1d31d51ad347601

Observation 1e56ca72-5632-47dd-a3c5-4f41127f93b4 · outbound

This paper cites Proceedings of the 25th International Conference on Artificial Intelligence and Statistics , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 25th International Conference on Artificial Intelligence and Statistics , series=

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.256433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.130420Z digest=sha256:633048a5f396137b1408049b67e7d346f6e4378b3b96ef84b823abe0afa42607

Observation 21bae2fe-2e8a-45f4-9c63-f696dd4eb6e2 · outbound

This paper cites Nature Machine Intelligence , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Nature Machine Intelligence , volume=

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.245235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.133891Z digest=sha256:b5524a2018a285de035dd4b37e54144422aacb2fb0cf51dffe5f0bef8a1eed51

Observation dda83449-15c7-4440-b343-65eef2a4f8dc · outbound

This paper cites IEEE Transactions on Computational Imaging , volume=.

Optimal Training-Time Scaling in Gradual Adaptation IEEE Transactions on Computational Imaging , volume=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.235261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.137163Z digest=sha256:a277ebef944d21f0593b66ae35683c80712325c76088ae2e1eb532b88de9f595

Observation ea11b1be-039e-4c82-ba3a-a108e88eb7e1 · outbound

This paper cites Advances in Neural Information Processing Systems Datasets and Benchmarks Track , year=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems Datasets and Benchmarks Track , year=

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.225014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.140355Z digest=sha256:13a028397e5dc4aa3ec54663784f762525909c41c3902b47c3d35e2275f06568

Observation 9713cd79-b8a4-4b70-a7d0-530f0c7f2efa · outbound

This paper cites Zico Kolter, and Ryan J.

Optimal Training-Time Scaling in Gradual Adaptation Zico Kolter, and Ryan J

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.214530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.143750Z digest=sha256:f680ed66730e9097e6911403068d790ea6d4ca2f470a5c30c99080ace075749c

Observation d8e4909d-b907-4bf2-a8f4-866881226438 · outbound

This paper cites an unresolved cited work.

Optimal Training-Time Scaling in Gradual Adaptation Unresolved cited work

Reference 57

Resolution
verified exact
doi, observed 2026-08-08T17:23:55.389672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.147041Z digest=sha256:007a90024cd2284f5a8dbf82e9be0b291f107f5e05b9db437cd2465718e19a1d

Observation 525b6b8a-5743-494d-8216-0e23aa91a675 · outbound

This paper cites A theory of learning from different domains.

Optimal Training-Time Scaling in Gradual Adaptation A theory of learning from different domains

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.150234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.150234Z digest=sha256:de85bc3893883d93f5de3ea67098aa75a26e013bf9006b7a399ae67ebcea9cb2

Observation bad29b34-21bf-45b2-866f-d0af9ecd2e5a · outbound

This paper cites Curriculum learning.

Optimal Training-Time Scaling in Gradual Adaptation Curriculum learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.153061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.153061Z digest=sha256:9689716d56cfec9d4c8626e8ee8226a739e0141dfe6a7888b5432b70ecabf97d

Observation 667f9263-4f93-4677-bd48-e071811041de · outbound

This paper cites Unifying importance based regularisation methods for continual learning.

Optimal Training-Time Scaling in Gradual Adaptation Unifying importance based regularisation methods for continual learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.204195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.155697Z digest=sha256:d48f0b9c92f9b780eef1987d4364434ce3d709cfb1d03f13fafc8121a4d6d962

Observation 32af7795-7e32-44c2-bf63-6c51dddddf4d · outbound

This paper cites Efficient lifelong learning with A-GEM.

Optimal Training-Time Scaling in Gradual Adaptation Efficient lifelong learning with A-GEM

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.192599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.158282Z digest=sha256:edef7c22e9847c71ac4b38c2936bdcc310e92066e49422bd7c9b017233916784

Observation ba1927b4-4334-4475-9e8d-e9b32a3d67d4 · outbound

This paper cites Gradual domain adaptation without indexed intermediate domains.

Optimal Training-Time Scaling in Gradual Adaptation Gradual domain adaptation without indexed intermediate domains

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.180838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.160809Z digest=sha256:94c80a363aa93dcbbf769b9174aac0ff035fb3210b886ffa48a48af3c584633f

Observation 9c8b1a8b-80bd-4395-88a3-5e0f785e7ee9 · outbound

This paper cites A continual learning survey: Defying forgetting in classification tasks.

Optimal Training-Time Scaling in Gradual Adaptation A continual learning survey: Defying forgetting in classification tasks

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.163358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.163358Z digest=sha256:cb4980d55935ae212c40257dc475926c2c9f6afc0e98d8ad32572834c15a02d0

Observation 3f2ca09f-6bde-4cda-ad4f-6d84f81fd830 · outbound

This paper cites Marsden, and Bin Yang.

Optimal Training-Time Scaling in Gradual Adaptation Marsden, and Bin Yang

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.167856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.165952Z digest=sha256:6c60b46b16d845e5fdb402eab75b1f4eb9640c41902a7ac3614651b215ac45da

Observation 19350aa3-5693-4370-83fe-412333536589 · outbound

This paper cites Dynamic update-to-data ratio: Minimizing world model overfitting.

Optimal Training-Time Scaling in Gradual Adaptation Dynamic update-to-data ratio: Minimizing world model overfitting

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.156451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.168575Z digest=sha256:101168065e958c5c0eab6745f760a0a948e20ed09ada30bd91f96887c3789210

Observation b737c49a-0958-4c88-9422-5a91fe48e588 · outbound

This paper cites Ward, Nathan Srebro, and Daniel Soudry.

Optimal Training-Time Scaling in Gradual Adaptation Ward, Nathan Srebro, and Daniel Soudry

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.145832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.171526Z digest=sha256:311f3cbe42e5b87d35e777e827766d1216565a8b9c5970d826423c33a5645909

Observation cf04c0e0-676f-4026-8cd9-695e5550a704 · outbound

This paper cites From continual learning to SGD and back: Better rates for continual linear models.

Optimal Training-Time Scaling in Gradual Adaptation From continual learning to SGD and back: Better rates for continual linear models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.133975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.174951Z digest=sha256:cc2f6d6d3c7666ef46e9fdee4135a5e10c4bfeebe01b904ac6292800ac08c6ac

Observation a0fd1e7a-1392-403b-9b9a-ff44a0e293ac · outbound

This paper cites Orthogonal gradient descent for continual learning.

Optimal Training-Time Scaling in Gradual Adaptation Orthogonal gradient descent for continual learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.122217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.178276Z digest=sha256:c9e7ea54048bd2d9f734756cdafe1bd46163dd2e644fef4a5e9abc5c6991b441

Observation 81540749-9d34-45e6-944f-989b88512639 · outbound

This paper cites Domain-adversarial training of neural networks.

Optimal Training-Time Scaling in Gradual Adaptation Domain-adversarial training of neural networks

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.181785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.181785Z digest=sha256:bbb205fb98d4703a5b413991e3c0e3ca761047b6739b769090ec1e073cf6a394

Observation b5025385-0bae-4ef9-ac6c-69f0d699f458 · outbound

This paper cites a henb \.

Optimal Training-Time Scaling in Gradual Adaptation a henb \

Reference 70

Resolution
metadata mismatch
raw_fallback, observed 2026-08-08T17:23:55.597503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.184959Z digest=sha256:2461025cf43a69f78e480617998d0bad1e674c2bbaf8368b8239947769e3eb3b

Observation 3242efd6-3052-4981-8000-e54763bf9572 · outbound

This paper cites The importance of being lazy: Scaling limits of continual learning.

Optimal Training-Time Scaling in Gradual Adaptation The importance of being lazy: Scaling limits of continual learning

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.102028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.188394Z digest=sha256:2286afba59a01e5c2a4512c608a7df7b1b1fb18a5686066346f1fdc5df1ff13d

Observation b9d84df5-0cf4-437c-9d25-2b2bbe59f688 · outbound

This paper cites Bellemare, Jacob Menick, R \'e mi Munos, and Koray Kavukcuoglu.

Optimal Training-Time Scaling in Gradual Adaptation Bellemare, Jacob Menick, R \'e mi Munos, and Koray Kavukcuoglu

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.089738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.191925Z digest=sha256:3093f791bc4d2f5356534ef6096608f55865ab410bbdf92b75a7b0091d8b5654

Observation a13ea8ad-30f6-451e-86ed-880454ec70ca · outbound

This paper cites On the power of curriculum learning in training deep networks.

Optimal Training-Time Scaling in Gradual Adaptation On the power of curriculum learning in training deep networks

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.077937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.195309Z digest=sha256:c296d08b49ec3cfb66f67dda01dd6a1a5ea0d7f82ca2788861e8e30b4acfaaa8

Observation 39c725d8-48fc-454a-9bc7-fd07d05d5fd2 · outbound

This paper cites Train faster, generalize better: Stability of stochastic gradient descent.

Optimal Training-Time Scaling in Gradual Adaptation Train faster, generalize better: Stability of stochastic gradient descent

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.066390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.199394Z digest=sha256:3410eee0bd72347d819420e64e8524680632449e1c852223ca45f681b1febbd7

Observation ecfdf522-f9ea-4d47-b8f4-888b1ecfbe11 · outbound

This paper cites Efros, and Trevor Darrell.

Optimal Training-Time Scaling in Gradual Adaptation Efros, and Trevor Darrell

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.054085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.202784Z digest=sha256:4336239af05e85cbe500d9c115a6728b4f9f858f28d638cd3246e3e45dd82d8d

Observation 07f56e1d-15a6-497f-a5b2-1541be0264f3 · outbound

This paper cites Curriculum reinforcement learning using optimal transport via gradual domain adaptation.

Optimal Training-Time Scaling in Gradual Adaptation Curriculum reinforcement learning using optimal transport via gradual domain adaptation

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.042421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.206339Z digest=sha256:dc96f0b538bffd7e2420dd9fb8af4d6c364fbf380047f99a43de2b3ff63ca8d6

Observation f980245a-77d2-4992-bc01-25e7c75483d8 · outbound

This paper cites On the adiabatic theorem of quantum mechanics.

Optimal Training-Time Scaling in Gradual Adaptation On the adiabatic theorem of quantum mechanics

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.209876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.209876Z digest=sha256:7465cf75f73c1747bc3d9b5957e99c9437532de8466287a48658d631f758a8f4

Observation 915bb5f1-f909-4f5d-a57d-65d7d5e9dbb8 · outbound

This paper cites Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell.

Optimal Training-Time Scaling in Gradual Adaptation Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.213423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.213423Z digest=sha256:d9b23b9d0e1f887b6a063c68b416019cb3f2d3fe1da6c392e252e7fb9c0e01d7

Observation 8efb8743-af9f-47cd-b4ae-70bd546a27aa · outbound

This paper cites Curriculum reinforcement learning via constrained optimal transport.

Optimal Training-Time Scaling in Gradual Adaptation Curriculum reinforcement learning via constrained optimal transport

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.029953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.216889Z digest=sha256:c1f93879eaaa90d61c1a998fed48b4705ed13ceef3332f2947851d5c6c4787e3

Observation fa6f26bf-c0aa-4ae0-b206-1b5a1c646498 · outbound

This paper cites Understanding self-training for gradual domain adaptation.

Optimal Training-Time Scaling in Gradual Adaptation Understanding self-training for gradual domain adaptation

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.017534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.220564Z digest=sha256:c19fb65a2ac7d226315ee0c54d98b6cc6f50934227f491dc77ce4e7cdb0f61c5

Observation 4fd5523c-07d7-4eaf-a037-4108aad70d6a · outbound

This paper cites Pawan Kumar, Benjamin Packer, and Daphne Koller.

Optimal Training-Time Scaling in Gradual Adaptation Pawan Kumar, Benjamin Packer, and Daphne Koller

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.004767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.224025Z digest=sha256:c75c4e56a1f8f0708b69d40aa819fbdebafb4ed8f71347d0c5b7ee2a1c7a3a56

Observation e768f1fe-6e5c-40ed-858a-d3baeb0be961 · outbound

This paper cites Challenging common assumptions about catastrophic forgetting and knowledge accumulation.

Optimal Training-Time Scaling in Gradual Adaptation Challenging common assumptions about catastrophic forgetting and knowledge accumulation

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.992349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.227650Z digest=sha256:57bf4b17268468fd9f5ac5504c0d689e953b2db1cff896f07235baacaa370578

Observation 17326707-639d-4958-899a-6990a7e19c3d · outbound

This paper cites Optimal rates in continual linear regression via increasing regularization.

Optimal Training-Time Scaling in Gradual Adaptation Optimal rates in continual linear regression via increasing regularization

Reference 83

Resolution
verified exact
raw_fallback, observed 2026-08-08T17:23:55.512645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.231073Z digest=sha256:ee146a7c43d73484f6429341987d231628882fc2638383b1a92145b00d51b088

Observation bc245f15-1dea-4e33-9dc2-d726e8e996c3 · outbound

This paper cites Gradient episodic memory for continual learning.

Optimal Training-Time Scaling in Gradual Adaptation Gradient episodic memory for continual learning

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.979861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.234563Z digest=sha256:0aaeb0f25c08d9c7d9ff06cdbf4a25044fd98a6bc948d4c49818509a499d017c

Observation 5758210b-7cf1-4e05-b5bf-1a0d2f393328 · outbound

This paper cites Understanding the role of training regimes in continual learning.

Optimal Training-Time Scaling in Gradual Adaptation Understanding the role of training regimes in continual learning

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.968182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.237802Z digest=sha256:4945441476ad3ee104a16e616855de76275f3e1f58460c83857b0b2a999d1f72

Observation d56dc019-7e13-4916-a1b9-0fc5d2e6b9de · outbound

This paper cites Optimal protocols for continual learning via statistical physics and control theory.

Optimal Training-Time Scaling in Gradual Adaptation Optimal protocols for continual learning via statistical physics and control theory

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.957586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.241255Z digest=sha256:8b5006bc3378e6f79aba747b3e9922daa2332cfa5018b2b9717601e2cbedf324

Observation 4c125bbc-2d11-4560-9ae1-8d9735659eec · outbound

This paper cites Efficient test-time model adaptation without forgetting.

Optimal Training-Time Scaling in Gradual Adaptation Efficient test-time model adaptation without forgetting

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.946293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.244578Z digest=sha256:8be87ab39542dc729408c695950d8a0165a9e74a84369b8853fbaa631e560201

Observation 9024b7fd-7cfb-4e77-bffb-8f0dee92606a · outbound

This paper cites Towards stable test-time adaptation in dynamic wild world.

Optimal Training-Time Scaling in Gradual Adaptation Towards stable test-time adaptation in dynamic wild world

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.935798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.247976Z digest=sha256:e3f70d55d8312b4b62e12af76edad291721acdc67af924ae1bbb2d9b08d0d691

Observation caa34e78-1ca6-4068-9c94-08a7db404630 · outbound

This paper cites Parisi, Ronald Kemker, Jose L.

Optimal Training-Time Scaling in Gradual Adaptation Parisi, Ronald Kemker, Jose L

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.251415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.251415Z digest=sha256:b79c4f1fffe1be04cc787672a607784b8ad54596125d2e178e5b853fdde7fd10

Observation b67db30f-b1f7-4a22-aeaf-11f79d86a641 · outbound

This paper cites Wainwright, and Bin Yu.

Optimal Training-Time Scaling in Gradual Adaptation Wainwright, and Bin Yu

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.923543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.255359Z digest=sha256:7176c810c2ef357f271df3c94b6bf5b7935f4df23e5823dfc8e31e5dbe7dfd0a

Observation faae7793-d04e-4acb-80ce-fd7d74b5252b · outbound

This paper cites Online structured laplace approximations for overcoming catastrophic forgetting.

Optimal Training-Time Scaling in Gradual Adaptation Online structured laplace approximations for overcoming catastrophic forgetting

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.911521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.258513Z digest=sha256:d1950c5443e0b1afbbc4a96e4923db826b9edbae5ae2edd894a6259635f6f016

Observation 3308c1a8-2a0b-4782-8e7b-a878b8be4ffe · outbound

This paper cites Maximum classifier discrepancy for unsupervised domain adaptation.

Optimal Training-Time Scaling in Gradual Adaptation Maximum classifier discrepancy for unsupervised domain adaptation

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.899814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.261834Z digest=sha256:545203737d7370bea0112c441364402936d18697cedd3969482a40feaf41bd65

Observation 5c8119ec-326f-4eed-a9f3-daba2c567ee2 · outbound

This paper cites Test-time Adaptation in the Dynamic World with Compound Domain Knowledge Management.

Optimal Training-Time Scaling in Gradual Adaptation Test-time Adaptation in the Dynamic World with Compound Domain Knowledge Management

Reference 93

Resolution
verified exact
local_arxiv, observed 2026-08-08T17:23:55.407862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.265348Z digest=sha256:95357e89c91ad9256a259343cc6ef0197826e01844bc3ba323aba608705aa1aa

Observation e85034c2-b5d1-4039-b690-45b880216757 · outbound

This paper cites Efros, and Moritz Hardt.

Optimal Training-Time Scaling in Gradual Adaptation Efros, and Moritz Hardt

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.887877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.269323Z digest=sha256:f6eeba8383c5990d796bc604b08c90398d93a8e4095405f7d96cf121466fcc9a

Observation bcf1c10a-402f-4a9c-800b-24b2f3135d74 · outbound

This paper cites Ward, Mark Kong, and Halyun Jeong.

Optimal Training-Time Scaling in Gradual Adaptation Ward, Mark Kong, and Halyun Jeong

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.875838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.272837Z digest=sha256:9cde3bd74ac1848d8a5972520348a1328f0349278f5baa598450a5c670b764dc

Observation 26f07ae3-e241-4874-b052-b950743b43b7 · outbound

This paper cites van de Ven, Tinne Tuytelaars, and Andreas S.

Optimal Training-Time Scaling in Gradual Adaptation van de Ven, Tinne Tuytelaars, and Andreas S

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.276142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.276142Z digest=sha256:ecb87db0962773b86317d0755e988e970dff23b5bf9d79f276453451f8f2ff0c

Observation a33c56d7-0b13-40a7-8a33-66bafc677b7c · outbound

This paper cites Tent : Fully test-time adaptation by entropy minimization.

Optimal Training-Time Scaling in Gradual Adaptation Tent : Fully test-time adaptation by entropy minimization

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.864000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.278984Z digest=sha256:8e253b0928519243fddf62b334243ebdd4c6910156c72085b2d43cd91c494502

Observation 10bad797-3a3c-4d99-8dee-aeef6e47b5be · outbound

This paper cites Understanding gradual domain adaptation: Improved analysis, optimal path and beyond.

Optimal Training-Time Scaling in Gradual Adaptation Understanding gradual domain adaptation: Improved analysis, optimal path and beyond

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.852148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.282327Z digest=sha256:2a7b12eff3f36cd9ac71388b2d1172043b68a4972f28f4cd97ff2e311393fd72

Observation 11742923-6c24-40fd-85ab-99d85f0cfec5 · outbound

This paper cites Continual test-time domain adaptation.

Optimal Training-Time Scaling in Gradual Adaptation Continual test-time domain adaptation

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.840391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.285133Z digest=sha256:f6aa9011933f12056d625852ee6513ab5d14cc25da4d761520a92c0b9f841e37

Observation fb8385d6-3e23-422f-9920-78cd65e8e0d7 · outbound

This paper cites Curriculum learning by transfer learning: Theory and experiments with deep networks.

Optimal Training-Time Scaling in Gradual Adaptation Curriculum learning by transfer learning: Theory and experiments with deep networks

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.828939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.288069Z digest=sha256:4873dc607cd7f0a85544efcd89bb2f61bdffb8d02e57cfee4a49903c58ae2dd5

Observation 1d3e465a-9131-4ccf-b05b-63a3a0103293 · outbound

This paper cites From Order to Distribution: A Spectral Characterization of Forgetting in Continual Learning.

Optimal Training-Time Scaling in Gradual Adaptation From Order to Distribution: A Spectral Characterization of Forgetting in Continual Learning

Reference 101

Resolution
verified exact
local_arxiv, observed 2026-08-08T17:23:55.748087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.291013Z digest=sha256:6bb47582633eb33a7f018a67e499656be72083108117ddf6e102e11cc0f1fe72

Pith citing papers

No inbound Pith citation observations are available.