Pith. sign in

Paper Citation Record · LEDGER

Reinforced Context Order Recovery for Adaptive Reasoning and Planning

As of 18 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 2 inbound Pith citation observations for arXiv:2508.13070.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.13070 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:23:21.085081Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T13:20:05.377960Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T13:20:05.555802Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact2
  • verified fuzzy12
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7469ccae-3677-4bb7-87b5-911ffedba46c · outbound

This paper cites GPT-4 Technical Report.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.320451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.320451Z digest=sha256:4bb03c6870c29bf181757115f6a2ef723cb4df04e936dcfc85e5ea356d48425b

Observation 84f7b816-81c1-43d3-a1cd-0561871b180e · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.325984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.325984Z digest=sha256:87a7d3535dd84c1d0078ca279997357401ded4fd3fb9960fe4f20e4e3d377f9c

Observation 86c64d48-c37f-4339-b76d-8c5cfd9b8312 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Gemini: A Family of Highly Capable Multimodal Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.331158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.331158Z digest=sha256:45c666f359d6820b6a6458b94cbb625619360a11e8a5ea162ad473ca340b26d4

Observation 298d0bbc-b325-420f-87da-ef90f1186c25 · outbound

This paper cites Argmax flows and multinomial diffusion: Learning categorical distributions.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Argmax flows and multinomial diffusion: Learning categorical distributions

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.336717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.336717Z digest=sha256:0ec02fdcd8714ab724d34f7961b822d76af6f234362a64b8105d37b19aa7a494

Observation c905a289-a497-4ea0-80a7-5875453e0f27 · outbound

This paper cites Structured denoising diffusion models in discrete state-spaces.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Structured denoising diffusion models in discrete state-spaces

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.342151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.342151Z digest=sha256:59713e5638f5f39472132eccc041a48ab3a159793daa952a46c0eed5dd5c86bc

Observation d13b7f15-35ac-4834-b368-1e01708ef52f · outbound

This paper cites Large Language Diffusion Models.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Large Language Diffusion Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.348659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.348659Z digest=sha256:1349ac63e0bbc149878e8cfda562158f695bd536a76af560769ea1bfd27fe3c6

Observation e47bdb06-d1c3-4c0e-a73e-51ef882f48e5 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Chain-of-thought prompting elicits reasoning in large language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.399436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.399436Z digest=sha256:dd6e00147e474f1c9906cdbeaba9582fd360ad6d1a8d40d41b1df2618246ef62

Observation 01175fc7-6e22-46e9-9b34-17927bc7b5c0 · outbound

This paper cites A survey on large language model based autonomous agents.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning A survey on large language model based autonomous agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.444145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.444145Z digest=sha256:029205eae870cb83845ec53b6a65a1086196a9ebbb360b068c883abf27f42fe3

Observation b09cb030-f1f0-4d84-a133-f51039183d36 · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.448688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.448688Z digest=sha256:0c8a650ddef9bb247eafdc3aeadbe5cdd079f255d7e695e3f025b5795e470b13

Observation 54bf7b2c-4eb5-43b7-8411-c679b471634c · outbound

This paper cites Causal language modeling can elicit search and reasoning capabilities on logic puzzles.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Causal language modeling can elicit search and reasoning capabilities on logic puzzles

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:22.395704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.453606Z digest=sha256:f6e704d00431ace5bc2c7224f858b792d793033dcde57d1aef445622f64b93ba

Observation 965622e3-4e8f-4ce2-a5fc-254e3f6b7f85 · outbound

This paper cites Beyond autoregression: Discrete diffusion for complex reasoning and planning.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Beyond autoregression: Discrete diffusion for complex reasoning and planning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:22.375745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.459325Z digest=sha256:5db38b1de854ffdab62d3a1b14b98d1a2fcfcb2224be6f3f4361ffd7d59d8f9a

Observation 05cf2003-d321-41ff-8fe1-09d15a573e3f · outbound

This paper cites Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked Diffusions.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked Diffusions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.464947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.464947Z digest=sha256:03358b8b6cf8d7335d286e5291b62a110909768b31475a50a2d8e1112e29bb92

Observation 9b497e3b-73ed-4ad7-a8e6-244403d68f4c · outbound

This paper cites Diffusion models: A comprehensive survey of methods and applications.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Diffusion models: A comprehensive survey of methods and applications

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.470797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.470797Z digest=sha256:a73d75e654e6305f1ff19d7caf592072c5cadc446a0dc374b89f1ae7dad65316

Observation 9a211638-26ef-4dbb-8f77-6d104cb1f768 · outbound

This paper cites A survey on generative diffusion models.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning A survey on generative diffusion models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:22.315211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.475682Z digest=sha256:e745267c51e006e59f0351eee60298ba35bff0ffd28df851718bb0e6f016cf2c

Observation 5ccadcb1-b9c3-4187-8069-e6163e26585e · outbound

This paper cites Autoregressive Models in Vision: A Survey.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Autoregressive Models in Vision: A Survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.480169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.480169Z digest=sha256:bf78a7a475cea8ceb6800673e8ba939c564f3b42a031107628ab149c860d96bd

Observation abdb9721-fe84-40ae-b71d-0d37a25bb70e · outbound

This paper cites Language models are unsupervised multitask learners.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Language models are unsupervised multitask learners

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.484472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.484472Z digest=sha256:dc66c6c3c5f95e221534127f668a13435abf3447ee6a70230319caac4486720b

Observation be4b17bb-5f69-4c0c-9cd8-52ac8caa2388 · outbound

This paper cites Language models are few-shot learners.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Language models are few-shot learners

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.488934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.488934Z digest=sha256:be8a083fc6b45797a33fb47431c6a9f59f3469f06adf790c0ef1d37d9afa4bad

Observation d0475881-861e-4b2e-9029-8b0282ac115a · outbound

This paper cites Denoising diffusion probabilistic models.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Denoising diffusion probabilistic models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.536615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.536615Z digest=sha256:354f5ba2cec9c03b446bb9c93ea0f76c3d1e18864b9ada7c7764d20df6067620

Observation 4f1f215f-056e-4198-b523-ae0bddb910a2 · outbound

This paper cites Generative modeling by estimating gradients of the data distribution.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Generative modeling by estimating gradients of the data distribution

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.623190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.623190Z digest=sha256:d283a77d4a2452026c96e44662c60a9816b6e398b3ed6750c9e6c62b2107eb45

Observation 20286b7d-3b56-47e5-b2ea-dce477f2527b · outbound

This paper cites Score-based generative modeling through stochastic differential equations.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Score-based generative modeling through stochastic differential equations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.627154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.627154Z digest=sha256:d75544e23e2b295cdabbfd96baa08f7464e96f168d9cde119aff9e40ab714c61

Observation fbb6b1e1-0dbd-40c8-be9d-d503b23bb464 · outbound

This paper cites Xlnet: Generalized autoregressive pretraining for language understanding.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Xlnet: Generalized autoregressive pretraining for language understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.631845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.631845Z digest=sha256:1483eb21da09856eb1126faba2ed64a3b976ff4b7d6199ea842828ec49593291

Observation dc348384-e46b-41f0-8dc1-baeae72b1fed · outbound

This paper cites Training and inference on any-order autoregres- sive models the right way.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Training and inference on any-order autoregres- sive models the right way

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:22.132496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.636751Z digest=sha256:db892042d53f0d644002a582d247c63a239a54da37b6108331265524da5b9b0e

Observation 8b4bec0d-bd87-41f6-a420-cd0c046bd30d · outbound

This paper cites Faith and fate: Limits of transformers on compositionality.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Faith and fate: Limits of transformers on compositionality

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.641147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.641147Z digest=sha256:8a56e9b0b69d44912b2cedd71666514c3b739f31a1bc489580945c0ab7df12a1

Observation efa3cb39-95eb-419e-a77a-36cc8eb1636d · outbound

This paper cites The pitfalls of next-token prediction.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning The pitfalls of next-token prediction

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:22.076460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.645817Z digest=sha256:75a4ece21f4c8e0fc44a04218652a8f56139afc487fe8d65b0c5e9e93dedacae

Observation 74239063-28d0-4d2d-a627-0f40d31a1492 · outbound

This paper cites Autoregressive Modeling with Lookahead Attention.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Autoregressive Modeling with Lookahead Attention

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:23:21.396791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.650030Z digest=sha256:a03c5d5d49a22e607d868d4af10978ae72fe534fc00359d1b98c03987bb421b1

Observation da8841ae-d26d-408e-bf5b-0a81161f3fce · outbound

This paper cites On the planning abilities of large language models-a critical investigation.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning On the planning abilities of large language models-a critical investigation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:21.915113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.654527Z digest=sha256:1dfa731099121235636790ed9b07a117caf3bce668d70ff8bd0067a46eed3f73

Observation e6d8d945-1483-4d39-8439-bd2e859162b9 · outbound

This paper cites Position: LLMs can’t plan, but can help planning in LLM-modulo frameworks.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Position: LLMs can’t plan, but can help planning in LLM-modulo frameworks

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:21.848871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.751591Z digest=sha256:c0231aee0adb5181a646a1aae6eb848f9ba099fa00c6f0ae3c2e5b63f99b24af

Observation 3521d35e-cb3a-40ac-91ac-257051d4cad7 · outbound

This paper cites Chain of thought empowers transformers to solve inherently serial problems.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Chain of thought empowers transformers to solve inherently serial problems

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.876473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.876473Z digest=sha256:cf7748c362dc0c01d77aee65a617f46d692a9b6a6937e7d81b576676eaef82eb

Observation 8efeed51-41c7-433a-be23-f0476711b8ad · outbound

This paper cites Discrete diffusion modeling by estimating the ratios of the data distribution.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Discrete diffusion modeling by estimating the ratios of the data distribution

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.908203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.908203Z digest=sha256:9d83226cf506cbd36c92f49cfcbbbcb68e410778848391d7a02e6a05a54c73be

Observation 25fa6405-107a-44ac-a6a3-7aac8e209cd6 · outbound

This paper cites A theory of usable information under computational constraints.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning A theory of usable information under computational constraints

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:21.805862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.913133Z digest=sha256:9a3388167a11740175c76eacff90a062aba46ce1a57486d1f00511c08e6bbfdf

Observation 273a925d-a8fb-467e-8026-0f063e74dfd3 · outbound

This paper cites On the shortest arborescence of a directed graph.Scientia Sinica, 14:1396–1400, 1965.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning On the shortest arborescence of a directed graph.Scientia Sinica, 14:1396–1400, 1965

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:21.789977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.917949Z digest=sha256:01bcfaf5064db4a0d75a7973158c0978c20b22f317f24e9436c2568a075f52e2

Observation dcf4f9d5-7aa3-4bbe-8f44-ad6d3b4e0ab2 · outbound

This paper cites Reinforcement learning: An introduction.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Reinforcement learning: An introduction

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.922686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.922686Z digest=sha256:9a08b42ba20996c06ab2f0c9edb0b0c3d45d44145c7b4dc507eea152a6de7dcd

Observation 1083529e-3d01-4309-a996-3368cd70d481 · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Reinforcement learning with deep energy-based policies

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.926708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.926708Z digest=sha256:59d5b560c20e4a7c7d2e4f0ff871db4aa843ea8eb34c3bb9a8b3e3b8117cbc26

Observation af1b85f7-8cc2-437b-b162-a22ae9d4dfd7 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Proximal Policy Optimization Algorithms

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.931444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.931444Z digest=sha256:fa400b02575a9f949ad5d170c069d2360f58508268e0386bd6abcd1c5fe38425

Observation e29a70e9-5aa6-44de-869d-cde724d26a9b · outbound

This paper cites Lee, Kangwook Lee, and Dimitris Papailiopoulos.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Lee, Kangwook Lee, and Dimitris Papailiopoulos

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:21.650947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:20.936418Z digest=sha256:9547b3e6352225dbe6ad128b6aa40e9a1a4fa10e5e667493061156ef83f49cec

Observation 84f0489b-cafb-4187-8ced-3fe79644a78e · outbound

This paper cites Positional Description Matters for Transformers Arithmetic.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Positional Description Matters for Transformers Arithmetic

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.940431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.940431Z digest=sha256:746becc1bc28fa3090757e0921045ce043cb9d425d8318bbb76468c31043172a

Observation 8d965c38-77da-454e-8428-b4256bb5d65f · outbound

This paper cites Reverse That Number! Decoding Order Matters in Arithmetic Learning.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Reverse That Number! Decoding Order Matters in Arithmetic Learning

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:23:21.195217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:21.046550Z digest=sha256:e7d2ac256c6712232cd9da30be037563addb44d514bb126b304b844abfa5a4f6

Observation 2d4ad7ce-dee2-41de-add0-9c8c7bb0538a · outbound

This paper cites Transformers can do arithmetic with the right embeddings.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Transformers can do arithmetic with the right embeddings

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:21.636104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:21.073895Z digest=sha256:40e95f075ec391c22b372cfba23ca431d096d2f6ed6c454e43392a81a9c98157

Observation 5f748538-feec-4615-9928-e9c37290ccb1 · outbound

This paper cites A Closer Look at Invalid Action Masking in Policy Gradient Algorithms.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning A Closer Look at Invalid Action Masking in Policy Gradient Algorithms

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:21.079193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:21.079193Z digest=sha256:7d0c0f5a30106abba7eeb6405c4aae8e99c00c4828d8b4f1dd175019f7cbfb4e

Observation 0b5c6b90-a96b-41d2-91ab-c5a11eb91d7b · outbound

This paper cites Focal loss for dense object detection.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Focal loss for dense object detection

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:23:21.618707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:23:21.085081Z digest=sha256:90151596cc4f8e46c9cb6b6f7e4f1cdeaa629a19c6bd860729f71b283e48b0cd

Pith citing papers

Observation 81703caa-0f45-45cb-aeae-62c064ed6958 · inbound

Any-Order Flexible Length Masked Diffusion cites this paper.

Any-Order Flexible Length Masked Diffusion Reinforced Context Order Recovery for Adaptive Reasoning and Planning

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-05T13:20:05.561020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T13:20:05.377960Z digest=sha256:d4224e9336bdbb35ced3f4ad3b4d12d64c485240788cbf54394046a8ea75bb7a

Observation 74d45714-c8de-4082-b695-fa2c9fc97465 · inbound

From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models cites this paper.

From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models Reinforced Context Order Recovery for Adaptive Reasoning and Planning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T14:25:38.146508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:25:38.146508Z digest=sha256:17a460de79b9deac5993c2f1939db7ad07220cc68a575e51b201af83d421b948