Pith. sign in

Paper Citation Record · LEDGER

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective

As of 20 August 2026, this Paper Citation Record lists 100 of 133 outbound references and 1 inbound Pith citation observation for arXiv:2505.19815.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19815 v1

Coverage vector

measured 100 of 133 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:13:10.322843Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:31:25.047058Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T15:31:26.249343Z

Reference resolution

100 of 133 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved86
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cd7960cf-27bb-447b-9b92-a329d3c9a498 · outbound

This paper cites On sensitivity of meta-learning to support data.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective On sensitivity of meta-learning to support data

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:03.846154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:03.846154Z digest=sha256:2aa53febd72c826cff65a1f15950c394ec7d880c992bf4e3df67f5ee123f769f

Observation 9062d4bc-6448-4df5-88c6-5edc78f9f64c · outbound

This paper cites Back to basics: Revisiting reinforce-style optimization for learning from human feedback in llms.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Back to basics: Revisiting reinforce-style optimization for learning from human feedback in llms

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:03.889689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:03.889689Z digest=sha256:bed9e62f2b35f57aabdf230ec1d5c2cf7268cb34d0719e774a1a89cb6d5a9a09

Observation 41bcb0e9-8fef-4824-a591-e2b94b7ae97d · outbound

This paper cites What learning algorithm is in-context learning? investigations with linear models.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective What learning algorithm is in-context learning? investigations with linear models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:03.954359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:03.954359Z digest=sha256:d11e2e909f5bebac530801ff7e49763705c6ffc4051d8d61557135c111868cf5

Observation 37cd0d8e-c446-4c34-9e67-58bc0e03c546 · outbound

This paper cites Hoffman, David Pfau, Tom Schaul, and Nando de Freitas.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Hoffman, David Pfau, Tom Schaul, and Nando de Freitas

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.007965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.007965Z digest=sha256:a81ca56aea3af35f9c99763f7affca2f6b2ffa432577ab94f468b0230d89743b

Observation 4d3070e3-e1cf-401e-9067-fef6d050f215 · outbound

This paper cites Transformers as statisticians: Provable in-context learning with in-context algorithm selection.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Transformers as statisticians: Provable in-context learning with in-context algorithm selection

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.021073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.021073Z digest=sha256:2cf7f8f7648df2e6d0b9bbee136e7d699ba7047e41d75e1207e9829e53b86ef3

Observation 640cc635-1292-4979-a2e8-f575c7d052ef · outbound

This paper cites Numinamath 72b cot.https://huggingface.co/ AI-MO/NuminaMath-72B-CoT, 2024.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Numinamath 72b cot.https://huggingface.co/ AI-MO/NuminaMath-72B-CoT, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.089527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.089527Z digest=sha256:9d728fc117d40d0e23f24edaf604a73f9ba568bdf51d0ce0adedab0a68fac640

Observation 6cbcf43e-f241-4849-8d4d-3a83eb7b4a52 · outbound

This paper cites On the optimization of a synaptic learning rule.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective On the optimization of a synaptic learning rule

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.153587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.153587Z digest=sha256:019c99e2d2826c64c940ac3e841f784ec19d6915bda8b1b267bc2770bd27813d

Observation 95fbf38b-d253-4a4d-9a88-36061aad59d2 · outbound

This paper cites Llama-Nemotron: Efficient reasoning models.CoRR, abs/2505.00949, 2025.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Llama-Nemotron: Efficient reasoning models.CoRR, abs/2505.00949, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.185963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.185963Z digest=sha256:de1f035a2fbc8d035aec3d43aaf9205a08825da48b4927d2c05356061256194b

Observation 63ffba2c-47ae-4b84-b451-cb57d72f734b · outbound

This paper cites In-Context Learning with Long-Context Models: An In-Depth Exploration.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective In-Context Learning with Long-Context Models: An In-Depth Exploration

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.222611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.222611Z digest=sha256:3511855821818e3305b8c114461b624ffec479f7b01d43057dae1e61e5757933

Observation 492d5d87-8e50-42d7-801b-64ae44e07137 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Graph of thoughts: Solving elaborate problems with large language models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.276040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.276040Z digest=sha256:93ec9e96f068838803b12ff6e6531e36c6ecfd90a5301a7f34fed12854f83323

Observation 22663e17-7e8c-4082-b7e7-5f4b984cf4e6 · outbound

This paper cites On the ability and limitations of transformers to recognize formal languages.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective On the ability and limitations of transformers to recognize formal languages

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.332975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.332975Z digest=sha256:70897087d2ccf450a70a057446fee896a93066ba53a54d5eb570056be133e611

Observation cbcbcbea-674a-4073-9ce1-ae6ee9130e25 · outbound

This paper cites Application of calculus of matrices to method of least squares: with special reference to geodetic calculations.Trans.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Application of calculus of matrices to method of least squares: with special reference to geodetic calculations.Trans

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.393726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.393726Z digest=sha256:ba5189971ef4eb8aaef59ef3109baa5cead12c6b8b5cc5232b53acdd90dc7577

Observation 0c356d39-3593-4cb4-96da-5e231a942a6d · outbound

This paper cites an unresolved cited work.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.463619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.463619Z digest=sha256:6ebefb05fb249e10181d274b9453a1e542dd934d9c64f82541ea6d93e883f386

Observation 111a5e79-786e-45ab-b28e-e508b1c541de · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.572965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.572965Z digest=sha256:25b7810e1d8188596d3d0ad3cd0b598df2d17147fbbd8412af5d2d601610867d

Observation 1d5550d9-631f-4a81-8b12-9f17cb116a5c · outbound

This paper cites A closer look at the training strategy for modern meta-learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective A closer look at the training strategy for modern meta-learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.650151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.650151Z digest=sha256:0d1d93057be30b7e67d90450d5aee42ba066069654611cd5159606843776bbd6

Observation ee5eaa55-4e10-49fc-b2d9-a4a2f9773045 · outbound

This paper cites Variational metric scaling for metric-based meta-learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Variational metric scaling for metric-based meta-learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.777215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.777215Z digest=sha256:07ab15fe2f05254d71d1b6a21190f94b5cbbf584640b1ebdae18cb64b601846c

Observation bd803e36-d4ad-4d2e-908b-1d81c32c6da5 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Evaluating Large Language Models Trained on Code

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.861032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.861032Z digest=sha256:5ba8538e8dec4dc3ae556a07afda3fb87b57ba2e02c584ae2abaaebc897850de

Observation 518ca357-322d-446b-80b2-59141bdd4cc3 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:04.956661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:04.956661Z digest=sha256:19df3fd8bab73183bd13a1543e55f40e191e5f27c44da02bd28ae0f1e25ea80c

Observation e55d93cd-53dc-4e59-90ce-8d4dac890a6f · outbound

This paper cites Tighter bounds on the expressivity of transformer encoders.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Tighter bounds on the expressivity of transformer encoders

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.037177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.037177Z digest=sha256:1d274efe64fe2191f0a8f7eb6055b31cb7c666757e5a3dc6533251d2152d42d1

Observation 13f9710f-010f-4439-b2c3-3413663f1975 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.097364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.097364Z digest=sha256:43749a56280b8bc23345429f9605aed5cf190c0344941ce224f25a6ae71b2cc3

Observation f2bb6e67-9f19-457a-a385-9c22bdff631d · outbound

This paper cites Gpg: A simple and strong reinforcement learning baseline for model reasoning.CoRR, abs/2504.02546, 2025.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Gpg: A simple and strong reinforcement learning baseline for model reasoning.CoRR, abs/2504.02546, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.185556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.185556Z digest=sha256:46cf704834f6ef49974aba83a6951be86458491f4431141398518c98668b45af

Observation 873e7d08-0884-446a-af88-6381fd5c39d4 · outbound

This paper cites Task-robust model-agnostic meta-learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Task-robust model-agnostic meta-learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.251388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.251388Z digest=sha256:79eaaa696f07c62bf171b75eadfc01b680e7d7d0bf2692810e61ccfbfc1baa7a

Observation 546d6461-e0ef-43e0-a34f-77188d01d39f · outbound

This paper cites How does the task landscape affect MAML performance? InCoLLAs, volume 199 ofProceedings of Machine Learning Research, pages 23–59.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective How does the task landscape affect MAML performance? InCoLLAs, volume 199 ofProceedings of Machine Learning Research, pages 23–59

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.313549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.313549Z digest=sha256:de2c7a4e02a2d2453f0ab7581494e657ef869e669e8a902e60eaad830e0f61cb

Observation 5d538284-4aa0-4d77-b52a-a7c81927e3b0 · outbound

This paper cites Approximations by superpositions of a sigmoidal function.MCSS, 2:183–192, 1989.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Approximations by superpositions of a sigmoidal function.MCSS, 2:183–192, 1989

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.373031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.373031Z digest=sha256:bb6aebbd72729ba06417bf8d83a816c013f1022e552675ac662ce7bd60a65ca5

Observation b3cd39d0-bcdb-480f-9e5d-544dbf71c5a5 · outbound

This paper cites Why can gpt learn in-context? language models secretly perform gradient descent as meta-optimizers.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Why can gpt learn in-context? language models secretly perform gradient descent as meta-optimizers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.423564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.423564Z digest=sha256:30d4620159ee228cbb9cc34c715f625cb9bb721c3f3c1fd49e82ed1e1f641308

Observation 4061d508-f39e-446a-a831-e4e9dd5c0317 · outbound

This paper cites Gemini 2.5: Our most intelligent ai model.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Gemini 2.5: Our most intelligent ai model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.535773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.535773Z digest=sha256:e03a4cf8533dd1def2de74bb4ada3103f0215d0a7dac6d6dc81ac1d06a575c6d

Observation d825fd48-c173-4c7c-ba2c-9bc5abf04bed · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.581094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.581094Z digest=sha256:e8ce9e765b7d10060220da30515a80975afcb4ebaaa560f31cf71db0d674e85c

Observation 68cacb91-93c5-4dff-b88e-87ab20cc52c2 · outbound

This paper cites DeepSeek-V3 Technical Report.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective DeepSeek-V3 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.623446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.623446Z digest=sha256:2254960816547c4e7b56ce17bd63342290922e963deab101243bff4e05682dde

Observation 8a67dbd7-0b59-4cbd-841b-f33895d00586 · outbound

This paper cites Universal transformers.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Universal transformers

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.680234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.680234Z digest=sha256:7fb5071e098f8fe1f05e7bea9a30f552a047530c86371756cff4a8c2927d7a03

Observation 5b31e925-f738-47dd-817e-06d3ab34c297 · outbound

This paper cites BERT: pre-training of deep bidirectional transformers for language understanding.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective BERT: pre-training of deep bidirectional transformers for language understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.778916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.778916Z digest=sha256:27f9ab7bb9fdefe112504059d00999ed9588138f0135119c5c90b1326870d9ce

Observation 0785688b-7ca1-42f5-b648-87b70ce5f28c · outbound

This paper cites A survey on in-context learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective A survey on in-context learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.847209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.847209Z digest=sha256:9913d0a9e034d78d320c6f95526ad24f09a58fdb07a6a71de7c6ad80fcccfb96

Observation 221e7f49-d486-4dde-8594-c117654f1f23 · outbound

This paper cites The Llama 3 Herd of Models.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective The Llama 3 Herd of Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.932438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.932438Z digest=sha256:481b156a8e4184b0617f326a3743160c5c4bda72dfca19b1cc364b4ba0c8df35

Observation a0b5cb80-3051-48de-847f-972378cc0d86 · outbound

This paper cites Towards revealing the mystery behind chain of thought: A theoretical perspective.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Towards revealing the mystery behind chain of thought: A theoretical perspective

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:05.991035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:05.991035Z digest=sha256:b3a1389b649ce8ef614ef2f906f6ebb7df9285bb2692bfc18c59912f46b3d26f

Observation 08087437-dede-417f-ac4c-65901acb7763 · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Model-agnostic meta-learning for fast adaptation of deep networks

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.102481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.102481Z digest=sha256:9afef7dc2233f058621524652e6d6249795c4dea1c0c3bc85c84f16709598ad1

Observation f3180c5d-5d3c-4818-b70f-c94029497cde · outbound

This paper cites Transformers learn to achieve second-order convergence rates for in-context linear regression.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Transformers learn to achieve second-order convergence rates for in-context linear regression

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.154765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.154765Z digest=sha256:7ca9fa7c5659d41c52ec829be556247c279302ce8dd610cebd0f6f6f84ee944c

Observation 9eff5f2d-8af6-41d4-bebd-65c8686419a8 · outbound

This paper cites Reddi, Stefanie Jegelka, and Sanjiv Kumar.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Reddi, Stefanie Jegelka, and Sanjiv Kumar

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.196331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.196331Z digest=sha256:f4b57a13b7e04614d527459f57be4e6b276160985c4cfda97796802f2d2e87cd

Observation 624d798b-e489-4310-b1d5-dc2edf934cd9 · outbound

This paper cites Lee, and Dimitris Papail- iopoulos.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Lee, and Dimitris Papail- iopoulos

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.230812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.230812Z digest=sha256:02554f3786075ff3c35f217b7ded40dd91d72b6bc095fc14d7f8dcf62bba888d

Observation 4448d8e4-a2ab-484b-89d3-7a10ae08bff6 · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.287596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.287596Z digest=sha256:c348757329725f1b57cf558dfd84baa4f47ff5a2e8ef267da2ddff6553f18b1d

Observation ffc3e076-c28e-4cfc-9980-2f1346756a49 · outbound

This paper cites Measuring mathematical problem solving with the MATH dataset.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Measuring mathematical problem solving with the MATH dataset

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.392230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.392230Z digest=sha256:d899636af81fc4fc7d782228144b32cc2a9d38032bf4316c3bf2e3410e2ae9ea

Observation bd24f75d-82fa-4466-b347-72ef1b765fe2 · outbound

This paper cites an unresolved cited work.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.430607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.430607Z digest=sha256:0f26a45b7bbae56775a9145c4ae125630654b8a0297e579c1b63761c02aef592

Observation f21210e3-b3ed-4504-8e04-fc6881df648b · outbound

This paper cites Approximation capabilities of multilayer feedforward networks.Neural Networks, 4(2):251– 257, 1991.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Approximation capabilities of multilayer feedforward networks.Neural Networks, 4(2):251– 257, 1991

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.473455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.473455Z digest=sha256:15d012a16b059a8e935b75600e5b7967c5af2d7966e26b182993ed5f63029423

Observation 15de22bd-2ef9-4a81-90c9-01be65b74530 · outbound

This paper cites Hospedales, Antreas Antoniou, Paul Micaelli, and Amos J.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Hospedales, Antreas Antoniou, Paul Micaelli, and Amos J

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.538821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.538821Z digest=sha256:4ad57ffb91e989a03abde6727995266be8a78e1ade7b5a6b8a5577a1c5bffc79

Observation 68ed48f3-f99e-45db-a421-0b288074d9c9 · outbound

This paper cites Universal language model fine-tuning for text classification.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Universal language model fine-tuning for text classification

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.609290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.609290Z digest=sha256:e88d97bd7abec0412a0f836a7dd9ee4f6ce3aab0e32af7df24dc0f6fded056b6

Observation 49759538-da87-46cd-b2f8-683f3257d031 · outbound

This paper cites Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.674033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.674033Z digest=sha256:671a14fd10dd2ea64faa28d07cb21fe5caad5c36da6346c4d169fc6beece20ef

Observation 04a28cdf-1d77-4bec-8a79-869984e94306 · outbound

This paper cites Transformers Learn to Implement Multi-step Gradient Descent with Chain of Thought.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Transformers Learn to Implement Multi-step Gradient Descent with Chain of Thought

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.748375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.748375Z digest=sha256:1418e534b78694349051815f28b043d4779a317d564fcfcfc089b41691f672ad

Observation 2517825d-c34d-423a-abe8-5ff36dab6d15 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Qwen2.5-Coder Technical Report

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.798749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.798749Z digest=sha256:a04d8a7f739d81e9870e381e9c2ef35c7e553715fca34bc8973675833ed9b0a2

Observation 02f2b3f1-399c-4b62-ad26-f141f0219f78 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.850119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.850119Z digest=sha256:eaf893bf0b3e90665ac640d8f7ee8a5d4622c422c72eb9a53e98b1f49dd7ca53

Observation 90e9767b-ae4a-4976-bae1-1253314a8f9e · outbound

This paper cites Xu, Jun Araki, and Graham Neubig.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Xu, Jun Araki, and Graham Neubig

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.927332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.927332Z digest=sha256:f593e15b9f437c215ac9f68facf160f63293c4e56dc26b0255998c1878f0a5e4

Observation 9a023e01-3435-474c-ba5a-78295caafe5d · outbound

This paper cites Efficient memory management for large language model serving with pagedattention.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Efficient memory management for large language model serving with pagedattention

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:06.987382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:06.987382Z digest=sha256:081b6d567bd46b76290717405547b086bb9970a22201fefa014640d3e1a7384a

Observation 1e7ae17b-a6b3-48a8-a2fd-d54b5352348e · outbound

This paper cites Bespoke-stratos: The unreasonable effectiveness of reasoning distillation, 2025.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Bespoke-stratos: The unreasonable effectiveness of reasoning distillation, 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.041699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.041699Z digest=sha256:553ed53e11b8d2d856ca839fe9c380ba7d466a94e2f0312055639dd284bfdce6

Observation 89db988a-f9ec-4f09-871f-95a0574a4636 · outbound

This paper cites Rupam Mahmood, Shuicheng Yan, and Zhongwen Xu.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Rupam Mahmood, Shuicheng Yan, and Zhongwen Xu

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.097294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.097294Z digest=sha256:eeb7d30055206ea7aeb1558ffda75737e5bb2448d67c0054f1c066d290542650

Observation b726312e-9fc3-4c7f-beda-8d93305874fe · outbound

This paper cites Meta-learning with differentiable convex optimization.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Meta-learning with differentiable convex optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.176651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.176651Z digest=sha256:944885fa5613d421473ae2cdee14b4abc1fe8738854990ff1ec0efcb5a629971

Observation 06ecbfd4-822e-4de7-869c-928c16e769ea · outbound

This paper cites Visualizing the loss landscape of neural nets.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Visualizing the loss landscape of neural nets

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.242294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.242294Z digest=sha256:501e709e146e0b7426d2faded1ae262be34db422054e7c3b28416bf16dd2989d

Observation 76f61f26-2816-44e5-bd00-473c7774b1a9 · outbound

This paper cites Learning to optimize.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Learning to optimize

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.318755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.318755Z digest=sha256:8fec67d174b96d76937ef3bc62dd882c9c80d2b4347ea69727a798fe6eba49e6

Observation f998e8fe-af3c-4a9c-9893-69e6a9366d7d · outbound

This paper cites Learning to Optimize Neural Nets.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Learning to Optimize Neural Nets

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.379623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.379623Z digest=sha256:6a010d0b59dd6abaed6d46c0cb2050f84601a6a33ee78375b2d527aaeac69fc8

Observation 27e8b020-eb5b-4619-9341-d453c3383f62 · outbound

This paper cites A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.442289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.442289Z digest=sha256:7a26d4098e0a38d99f6fc31941284b8a4974071eb37090185ff70d5308fb7c21

Observation 6931b8cd-e0db-4d64-a1c0-2f141e2dfed9 · outbound

This paper cites Let’s verify step by step.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Let’s verify step by step

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.493931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.493931Z digest=sha256:dafeda9f9eb925ffeb4ea721304c7dd302770f9686d3f2a427ad2c54326c779e

Observation d629610e-a7cc-49a7-a1b6-d5468648bd1f · outbound

This paper cites Ash, Surbhi Goel, Akshay Krishnamurthy, and Cyril Zhang.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Ash, Surbhi Goel, Akshay Krishnamurthy, and Cyril Zhang

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.561262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.561262Z digest=sha256:ea46fa731195f79cc8ed3f65bd4e35c79c3c3d72cc6419356e872de51e661b6b

Observation 78b9157b-413b-488f-b7d7-5aea2ac871d9 · outbound

This paper cites an unresolved cited work.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.624447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.624447Z digest=sha256:5ed3654be7f7379c5576b735d1a0b3921f02c9bb43cc0e35c086a885f2c766d3

Observation 011da701-fbe3-44ef-af4e-b2ac622f280f · outbound

This paper cites Are Your LLMs Capable of Stable Reasoning?.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Are Your LLMs Capable of Stable Reasoning?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.676651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.676651Z digest=sha256:2e19f86e3017765ed56c7904fd8303d29223f169bc3a26b00110d60901ccf772

Observation fd37c512-1417-436e-851c-aa6767c0faa7 · outbound

This paper cites Reasoning Models Can Be Effective Without Thinking.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Reasoning Models Can Be Effective Without Thinking

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.755282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.755282Z digest=sha256:5b96a1210674b081c83611dd9c15cb4bce756c527022b9e40192ecafe68a64bc

Observation ee90d448-51da-4cb6-9e8a-5adcc25a3fb0 · outbound

This paper cites an unresolved cited work.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.821020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.821020Z digest=sha256:38977b214dd1af31f69384a87993f2262190e2440e4fc38df7152742eeb73d56

Observation e53b52d6-0857-47c1-9c75-47ac5611c292 · outbound

This paper cites Academy, 1877.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Academy, 1877

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.888575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.888575Z digest=sha256:99e9cab2705fcae9b7cb446ca87fa3f132816a43af60da525e95627d2e90c319

Observation 7fb5527f-58c4-4ac9-b746-3f9b15ecad3a · outbound

This paper cites Metaicl: Learning to learn in context.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Metaicl: Learning to learn in context

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:07.949373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:07.949373Z digest=sha256:c3eccafa8cf55436925616444a1e7f3163bdcd5296f6cc07eec568fa95aab6ab

Observation fa2536c2-34e6-46df-8164-eb0c6387edc8 · outbound

This paper cites Rethinking the role of demonstrations: What makes in-context learning work? InEMNLP, pages 11048–11064.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Rethinking the role of demonstrations: What makes in-context learning work? InEMNLP, pages 11048–11064

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.015925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.015925Z digest=sha256:010364f1a4d09187fadf81d17e485cc4dbaa751ce6da8a305e862b769af385ec

Observation 550dffcd-cfeb-4f98-970f-5d2d4c3956f4 · outbound

This paper cites Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.100248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.100248Z digest=sha256:d44c2cc2531e97b0bad1f520ad1cee1a050d11daabd422f40a23979642407a02

Observation 22b59dbf-ee13-4759-ad72-fd323d8caa7f · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Playing Atari with Deep Reinforcement Learning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.198094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.198094Z digest=sha256:0a3f3ae1645741ce600eea761891b7253c3499ced916f631a4dcacf2dd5c04f6

Observation bf72671a-9350-459f-bc18-2878a716ae4a · outbound

This paper cites On the reciprocal of the general algebraic matrix.Bull.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective On the reciprocal of the general algebraic matrix.Bull

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.267336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.267336Z digest=sha256:e239057a4798cc2a76277c93749f60c3e96ec48205818009e2a556ddeb610483

Observation 64fec554-2a0e-4197-b806-4203055af3a3 · outbound

This paper cites s1: Simple test-time scaling.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective s1: Simple test-time scaling

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.379653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.379653Z digest=sha256:ab314abd976c34db05a3f5bcf024950d6fbdca2369e27969e0edf4eb279fabf6

Observation 0f35de02-d51b-41f4-88f5-a971b38fd0e8 · outbound

This paper cites In-context Learning and Induction Heads.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective In-context Learning and Induction Heads

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.434201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.434201Z digest=sha256:07f6b68c2726870b09fbe746372a95879091e3f373f74214aec9fc7ca69db74e

Observation 9ee7c885-1c6d-4409-845d-bea84e13b2ae · outbound

This paper cites GPT-4 Technical Report.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective GPT-4 Technical Report

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.513001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.513001Z digest=sha256:ffb765dd31fcbadf35c174a658d0b77c57fc7f18d8e6ed79bd4a22a635257552

Observation 683dd9a5-4a8a-40ff-af3c-4ffd1827649e · outbound

This paper cites Learning to reason with llms.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Learning to reason with llms

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.575853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.575853Z digest=sha256:a5ecbab3d64ff95548aa3d65478041abc492e74e05c815b61e2ae04d158f7a21

Observation 9d3b5eb4-c868-4973-9009-be573e8d2175 · outbound

This paper cites an unresolved cited work.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Unresolved cited work

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.641941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.641941Z digest=sha256:7bf60c3bf8fa963945c3b911fe81556ac0c1fe3c00bdfd6ffbba27b0603e74cd

Observation f219487c-ed62-43de-98a5-51cc592c890f · outbound

This paper cites Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.705867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.705867Z digest=sha256:5fbfca1e98b0579932d67f2369e82c96764d39bc3fd5c109844c765951a52688

Observation 7e03045f-d25d-4ae7-bc7e-351a68c5b162 · outbound

This paper cites A generalized inverse for matrices.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective A generalized inverse for matrices

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.774117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.774117Z digest=sha256:dd4093cfcd083cc56124c547224663b5577a67e2303b9677aba1d8bb1b5b6124

Observation b5119c13-0e97-4ebe-83c0-3850e6a154b5 · outbound

This paper cites A simple guard for learned optimizers.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective A simple guard for learned optimizers

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:21.684679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:08.870312Z digest=sha256:f678396b6161650a493beb4be4c39dfe7470a95d10c69d4f2f1d56ac6850928f

Observation c47a79b1-30cf-4b43-a5c0-f515f443b4e2 · outbound

This paper cites O1 Replication Journey: A Strategic Progress Report -- Part 1.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.930615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.930615Z digest=sha256:4096086d21cbe9b31abc5d26987997498b33c26ec2133317b892a658f65f1d46

Observation c1011af8-99ce-4177-8156-c40f5efe5df3 · outbound

This paper cites A survey of efficient reasoning for large reasoning models: Language, multimodality, and beyond.CoRR, abs/2503.21614, 2025.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective A survey of efficient reasoning for large reasoning models: Language, multimodality, and beyond.CoRR, abs/2503.21614, 2025

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.992883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.992883Z digest=sha256:be792515deae28dc39314bdb87509146b1cef341f73f3e16d353d3f5de23a42a

Observation 2c8d8de0-ee38-49d5-8af1-33a0e84922af · outbound

This paper cites Improving language understand- ing by generative pre-training.OpenAI, 2018.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Improving language understand- ing by generative pre-training.OpenAI, 2018

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:21.531829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:09.058562Z digest=sha256:bf219d718bbc90396bc9d4f3c6988802e074c14408690f905f5b8e8bd3c28737

Observation e51bd8b2-4182-498d-96cc-3b7833cdff6d · outbound

This paper cites Manning, Stefano Ermon, and Chelsea Finn.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Manning, Stefano Ermon, and Chelsea Finn

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:21.300794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:09.133936Z digest=sha256:522ad638a3599b5438adbccc8c2bae165e1e68a8fc767dea8d3f8a5d00d1347d

Observation a034b155-5f00-40d9-9edd-e50bca4beed3 · outbound

This paper cites Kakade, and Sergey Levine.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Kakade, and Sergey Levine

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:21.122628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:09.195493Z digest=sha256:2b41b0a048197e1113222c3dc245cf5fdbfb4a1e04ee7d26946eca7c6f5595d5

Observation a8b5061b-8020-4e77-8b7f-93b94df87b7f · outbound

This paper cites Optimization as a model for few-shot learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Optimization as a model for few-shot learning

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:20.922435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:09.270372Z digest=sha256:1f2c890173e8d78ec06a9aebf38bbd2d8b1f0598562b5123d720d2657bac0cb5

Observation a65b5a9c-ba39-4bdf-aa26-e2de63b7b95d · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:09.318060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:09.318060Z digest=sha256:acfa911d2a9a13d023c4960f1fc892d8a3d387710bc14c550dfd729ee80f0038

Observation 5330e67a-6131-43f2-85f1-333ad8ba8a99 · outbound

This paper cites Learning to retrieve prompts for in-context learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Learning to retrieve prompts for in-context learning

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:20.774803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:09.365657Z digest=sha256:33e7f5a020aad45f0f70dfa70e06f057e8264b949fbb8fe9db003a0d69090461

Observation 93d8f48c-b70b-4b15-8978-319febdeeaaa · outbound

This paper cites One-shot Learning with Memory-Augmented Neural Networks.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective One-shot Learning with Memory-Augmented Neural Networks

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:09.429316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:09.429316Z digest=sha256:608078c8cfdd8edd216c3ac30d11ad601feb433051ef6229ad1c057a56fe5ef9

Observation 4fe45818-2782-473c-abbb-d63ea510437a · outbound

This paper cites Evolutionary principles in self-referential learning, or on learning how to learn: The meta-meta-.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Evolutionary principles in self-referential learning, or on learning how to learn: The meta-meta-

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:20.546970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:09.524137Z digest=sha256:f5f2d0d12ca20d54b2a114d2ee9eee88c1a06651e66e548667791745f3be4bc3

Observation 93df2217-f2c0-45f0-ab75-75ad274dafc7 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Proximal Policy Optimization Algorithms

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:09.625547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:09.625547Z digest=sha256:0fd090a1e8f5c84cdad39fc2f8758182dd9eb67b21f91bb5eb52b40fadbb8d85

Observation add5ac14-23fd-414a-84bc-edfcea2c2eae · outbound

This paper cites Rethinking Reflection in Pre-Training.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Rethinking Reflection in Pre-Training

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:09.701448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:09.701448Z digest=sha256:75860afed89ab9d1e88ef6d29c6684898f485d900f633f695a9b4e59b9142a90

Observation c3388108-bb18-4d55-a6ac-eb23367b7892 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:09.795838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:09.795838Z digest=sha256:e719e962d89ae1a1ae52ca43c66d8045e57dee7ff274ff10bf8deb68b545b0aa

Observation 52505d35-b8b5-4057-8449-f549907560a3 · outbound

This paper cites Hybridflow: A flexible and efficient RLHF framework.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Hybridflow: A flexible and efficient RLHF framework

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:20.357532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:09.853973Z digest=sha256:70498ed65442b48f36cff4019d77103a95fdb315fbf52aa9429dc14785e80aa6

Observation 9ade0988-831f-482c-85dc-b309ea41f64b · outbound

This paper cites Route Sparse Autoencoder to Interpret Large Language Models.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Route Sparse Autoencoder to Interpret Large Language Models

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:09.892396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:09.892396Z digest=sha256:cdc153277fb545b84042f2a1c0804bd30e0a49d1160a35b94259435983d50bde

Observation 6c376dd7-ff2d-409c-8abe-555878c0ed4d · outbound

This paper cites Co-Reyes, Rishabh Agarwal, Ankesh Anand, Piyush Patil, Xavier Garcia, Peter J.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Co-Reyes, Rishabh Agarwal, Ankesh Anand, Piyush Patil, Xavier Garcia, Peter J

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:20.145364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:09.944823Z digest=sha256:d037faafe556399fa2233196c32edbbd94616dedfc28812e1d85bee711efd3f6

Observation e8ad6c2e-f46e-4cf0-8cef-7ad7ff522a18 · outbound

This paper cites an unresolved cited work.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:13:19.926428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:10.004740Z digest=sha256:de42826e647ad1894ed13d9827f005ed2e4287da2a75d3038a61d01a5bf69de0

Observation 798131f2-a93a-48e7-9826-3de2136420aa · outbound

This paper cites End-to-end memory networks.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective End-to-end memory networks

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:19.761948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:10.050171Z digest=sha256:7ba95a01d25996dd01036761fb7db8b18961b4c917b051db0d31b919c8919907

Observation a9dd52e9-1da2-455e-b46e-8e90b913ea8c · outbound

This paper cites Andrew Bagnell.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Andrew Bagnell

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:10.096820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:10.096820Z digest=sha256:cf77504d06be220c15d3b5234aa988e36f9315c60bfbbe31f8b9b5a25ba48a47

Observation 1a7e9674-6b59-4031-8e4e-44638681b15f · outbound

This paper cites Blockmix: Meta regularization and self-calibrated inference for metric-based meta-learning.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Blockmix: Meta regularization and self-calibrated inference for metric-based meta-learning

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:19.584497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:10.141251Z digest=sha256:6b28afdd01386a483b6c0ceb8b3dbcda0b4b9078d25a10f7f716bcdd4259afd9

Observation 0dc475fb-517a-4ba8-9778-89e2b58fd9fe · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:10.188547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:10.188547Z digest=sha256:42807bd5a9aba3c560b14a2d8f93464d9b36b0eb76b3722deee9695a6c4407ef

Observation 0b24141c-efa8-4b5f-a8bf-a283380141c8 · outbound

This paper cites Sky-t1: Train your own o1 preview model within $450.https://novasky-ai.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Sky-t1: Train your own o1 preview model within $450.https://novasky-ai

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:19.414611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:10.233784Z digest=sha256:2ac1232991716ea90aedf5b61c3952bf31aaff14fe6a4ffc849bbfc6f58c5544

Observation 8d544828-9521-49e4-87a2-9fd7f3cb77b3 · outbound

This paper cites Open Thoughts.https://open-thoughts.ai, 2025.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Open Thoughts.https://open-thoughts.ai, 2025

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:19.181089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:10.278438Z digest=sha256:10dc3443b288eb580b39f2b31a3d5cd1d85f9fba7972c9f56fa58f158a41ab7b

Observation 1a42a299-166e-4879-9e37-ed8dfed4d730 · outbound

This paper cites Qwen3: Think deeper, act faster.https://qwenlm.github.io/blog/qwen3/, April.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Qwen3: Think deeper, act faster.https://qwenlm.github.io/blog/qwen3/, April

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:13:18.893561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T14:13:10.322843Z digest=sha256:632b58e327454f448706fed1fa36bd23f5e18ea74abd536f85835edfa46f43e6

Pith citing papers

Observation 36db4986-8f7d-4b3a-bcd8-080ad3e7dd9e · inbound

Learning without training: The implicit dynamics of in-context learning cites this paper.

Learning without training: The implicit dynamics of in-context learning Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:31:26.311146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:31:25.047058Z digest=sha256:fb7d604035d9041bb1d2ec2860d12875557c9a6cd1f5cc79d2ef7c84c6ef59ea