Pith. sign in

Paper Citation Record · LEDGER

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

As of 21 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 13 inbound Pith citation observations for arXiv:2506.13923.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13923 v2

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:31:38.586086Z

measured 71 of 71 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:19:46.624230Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:48:56.210532Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 93c69185-51db-403f-b91c-1b1ce5ffd1d1 · outbound

This paper cites OpenAI o1 System Card.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models OpenAI o1 System Card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:33.630703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:33.630703Z digest=sha256:121a79a19f73e05efba4738ae9ff64e2feebf3247dfd75982d44ad55c6eb4cc5

Observation a8988f1e-0c2e-4284-bd35-823af1320e74 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:33.673149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:33.673149Z digest=sha256:42134de1c6ce10f9604b4dbb7ac4497d94dcf29a275f85d58993f3d952fa90b3

Observation b7ba5a2e-91c5-42a3-9966-0dd733e5b83c · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:33.763546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:33.763546Z digest=sha256:5c042aa50f6d8e2d3d6acb9cae326410b5da68f10c9452217bcb8a27c350d1e4

Observation aeed9b1a-ca68-44e7-a2ec-7aa18495b3a8 · outbound

This paper cites Teaching large language models to reason with reinforcement learning.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Teaching large language models to reason with reinforcement learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:33.824990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:33.824990Z digest=sha256:2b403b0c00df96f565e9cae9b9a9ba5aa3f8e5e49a20d31b386f8944dfc7326d

Observation d0963d46-c208-49b6-a62c-dfbab7fe6557 · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:33.883507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:33.883507Z digest=sha256:cc61cdc911393641d2b3fc53d97775fbc0e3f705167ed786f2375d1e8ee4abd2

Observation 13e84987-c960-42bc-9126-4f976a084cc3 · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:42.968767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:33.976587Z digest=sha256:7507c0db6ad959fb9d139438d0e36ced6c088263c70cf9d994058a4c84fcd000

Observation e30394ea-1a98-4aed-b777-7263413883c9 · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Reinforced Self-Training (ReST) for Language Modeling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:34.071107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:34.071107Z digest=sha256:6baaec2dcaafbaa6e5a7c828227d099e7bc2cae261cea6f99aae3041071c7ca0

Observation 689b6931-b901-48bb-96f4-67616561938a · outbound

This paper cites Openai’s reinforcement fine-tuning research program, 2024.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Openai’s reinforcement fine-tuning research program, 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:42.742679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:34.138304Z digest=sha256:32c75b202f94db1ba8029c270a7e5821f8d7b8d6547363919c87bc31fd95a98c

Observation c1892acf-dbe5-43b7-95c3-e78b068ceaea · outbound

This paper cites Be- yond human data: Scaling self-training for problem-solving with language models.Transactions on Machine Learning Research, 2024.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Be- yond human data: Scaling self-training for problem-solving with language models.Transactions on Machine Learning Research, 2024

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:42.566202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:34.205487Z digest=sha256:46e3173a26b504146548e357036bf796970d8cfa508db33b778d38f9ccc0a00a

Observation 77cc6ec1-5c4b-4636-94df-9c093f623c76 · outbound

This paper cites V-star: Training verifiers for self-taught reasoners.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models V-star: Training verifiers for self-taught reasoners

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:42.501094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:34.323233Z digest=sha256:e912e3044d8753afdd96991619de907fe79d28735d3d38e14fb4d06337380c40

Observation c2bb22dc-bb45-4d6b-ab4c-61751a999089 · outbound

This paper cites ReST- MCTS*: LLM self-training via process reward guided tree search.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models ReST- MCTS*: LLM self-training via process reward guided tree search

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:42.266560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:34.410205Z digest=sha256:c0714f7d7531f22a59dab880a664fb12e61ec9f4972683e3de2a855a6eaef576

Observation b2a05105-dcec-47c0-9c68-e37529ec670d · outbound

This paper cites Continuous control with deep reinforcement learning.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Continuous control with deep reinforcement learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:34.508482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:34.508482Z digest=sha256:75b87b09546ad29a192ba8be246c873c31cb4a2ffceca3b7c777e0f0d0326acc

Observation b0c55e14-5657-40b7-af77-ff320e39af2b · outbound

This paper cites Qwen2.5 Technical Report.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Qwen2.5 Technical Report

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:34.635249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:34.635249Z digest=sha256:c99d719415f964cc1c1f83b581def5a2727805cbcee469eb2a38234d65260528

Observation 503ce4ba-3497-4e35-aff9-16f65a6e7112 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Measuring Mathematical Problem Solving With the MATH Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:34.851692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:34.851692Z digest=sha256:e5ab0b6e14f8b31e0c39ea39859c0ec59ee0f073cd525964e17debb57228cc5b

Observation a94a7577-dd41-45a5-befa-721617b21833 · outbound

This paper cites Aime 2024.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Aime 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:42.081687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:34.924828Z digest=sha256:ab871deb9e37e170ce5fe428261046a54592331fdacac188e3a4a7e8ec882191

Observation 27f7e066-ca93-4f55-bde1-e8f978f7cb64 · outbound

This paper cites Aime 2025.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Aime 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:42.010708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:35.021740Z digest=sha256:cb1840d95c81da5e54427930658df90173d8e44a7ca64923df2e1594232235e5

Observation 5a8cccb6-bcf6-462a-af42-818c0d348387 · outbound

This paper cites Amc 2023.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Amc 2023

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:41.885367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:35.145729Z digest=sha256:cd1b3870a404067cf23159071ead12ee96ea4b0b5952909f99dc90d35575155b

Observation 07087427-359a-4e9b-9f33-1c67cb2dfdfa · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Gpqa: A graduate-level google-proof q&a benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:35.213884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:35.213884Z digest=sha256:c7aca86d7abaff24ee42da8cf75011f15418caaa8192163ccc1f97f1a248e133

Observation c05f9dc9-279f-4538-b816-a2bec136b49b · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:35.273503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:35.273503Z digest=sha256:ed54d9e18246333e95cb16fbd3130ecfd188ec44ac927b9e9dcbf12b2e5fe4a4

Observation beb9fd64-2327-4d99-9065-6f3e2b8f317e · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:35.407856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:35.407856Z digest=sha256:eae1c5a87aef69d6ad7db58de99876d23eba5f60ec6bfddc0b31d0f100d16fd3

Observation 2577b0f8-e45f-4c7a-a99f-bf593af79254 · outbound

This paper cites Livecodebench: Holistic and contamination free evaluation of large language models for code.arXiv preprint, 2024.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Livecodebench: Holistic and contamination free evaluation of large language models for code.arXiv preprint, 2024

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:41.787072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:35.477463Z digest=sha256:54f6c2da47a75bf2e51f72526885237cfabf3d761458ef0203dca510acf9fbbb

Observation 098d9765-40b6-4420-900c-835487736344 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Evaluating Large Language Models Trained on Code

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:35.563471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:35.563471Z digest=sha256:712494fcfab6d3d074b6129accf6570d6ec23ce79fdb6bd97c161efdf48ed4cd

Observation 35b5e4f3-a3a6-487f-a86d-414cb3cac3d3 · outbound

This paper cites Open r1: A fully open reproduction of deepseek-r1, January 2025.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Open r1: A fully open reproduction of deepseek-r1, January 2025

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:35.647369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:35.647369Z digest=sha256:dbc8ecf5854ac18d98bf254ff1100b0f3d1c199417dde96b164c5e9ade7dbb41

Observation a9ff5d5f-0123-4e33-b672-8e7b2dd1c244 · outbound

This paper cites Dapo: An open-source llm reinforcement learning system at scale, 2025.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Dapo: An open-source llm reinforcement learning system at scale, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:35.732089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:35.732089Z digest=sha256:35f70514374e11b1ca009b45c9b36ce58366b5e325aa2e883b90720d69a5b283

Observation c7094ab0-9a3c-441e-b0b6-b7e5c9eb3c5e · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Fine-Tuning Language Models from Human Preferences

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:35.831960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:35.831960Z digest=sha256:c9286271af2fefb56f7de1fd4246a4cf9d4fe47e2430546f590784dd5288e5d6

Observation 97b7a885-7d76-403d-8754-9278b13936d2 · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:41.628336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:35.904313Z digest=sha256:45ae7cf9370b9227cb93623f6897bbf44398052d7b1e8eb8d48c1fbc3b9aadad

Observation a9c0a123-1458-4ec2-ad46-403029fe2937 · outbound

This paper cites LMRL Gym: Benchmarks for Multi-Turn Reinforcement Learning with Language Models.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models LMRL Gym: Benchmarks for Multi-Turn Reinforcement Learning with Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.028658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.028658Z digest=sha256:a8e988332ef8d9f3154aaf40703d1f4b61c10dce5cc53f4a041d863e86615139

Observation c63c58de-bbae-4044-892f-a21618e9f075 · outbound

This paper cites Archer: Training language model agents via hierarchical multi-turn rl.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Archer: Training language model agents via hierarchical multi-turn rl

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:41.511850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:36.112436Z digest=sha256:7ee616d7179dd42aeb3710a024f4e36f5e586ea095aca6c32ea8aa592f0d13f7

Observation 9a9491a7-4fca-4fb7-b1ee-481b3c08470d · outbound

This paper cites Aligning large multimodal models with factually augmented rlhf.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Aligning large multimodal models with factually augmented rlhf

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:41.386519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:36.165728Z digest=sha256:76a2083019af01ea23b6aa5912a86d5fe6b4854779f4f7a25b4ccc1f6eab61b4

Observation 63cc29ad-0a40-48b2-b947-c9ab58369e60 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.250869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.250869Z digest=sha256:0a1c1a2e97bb467998c5889f52581faaab68e5e090f0753e468f66755d7ee9d1

Observation 62593bd3-d18c-4b33-b0fb-2ab464d70a43 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Training Verifiers to Solve Math Word Problems

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.352872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.352872Z digest=sha256:00d6098af3ac39e7bee90c22f06e61587cb65a49f59686d9a5d3ffefd1c0e061

Observation a56e454b-536e-4593-a6ea-861ee4fba24f · outbound

This paper cites Program Synthesis with Large Language Models.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Program Synthesis with Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.490405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.490405Z digest=sha256:d7bd7173f5419cb8baa7e5c51061b253d6e2dd7faf05d33cddf7acf47af33416

Observation 43b36156-2531-46f0-b9dd-afe892cd943c · outbound

This paper cites RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.562765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.562765Z digest=sha256:1f03eef25c5257d59261ad40acec1d2376de75842d471616fc2074998af97d00

Observation 0e997cca-1852-4654-a328-b4b1db39ef8c · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.601300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.601300Z digest=sha256:7b3ed5df9044ec788712708444b2d4a81bc542c4cfab8a78d7d5f80847815578

Observation 386ef46f-11a8-45a7-9bcd-9bf7868f5487 · outbound

This paper cites Scaling LLM test-time com- pute optimally can be more effective than scaling parameters for reasoning.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Scaling LLM test-time com- pute optimally can be more effective than scaling parameters for reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.691691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.691691Z digest=sha256:4cab6bac6dff9bdf54999730a776563a864e39f6a73123eb2a911cd6cbecd0b1

Observation 3f71c7fa-5d4b-4ef8-a912-a461dd82f61c · outbound

This paper cites Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.795096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.795096Z digest=sha256:ea896d3a26d0bb13db7b7281cb7ebfc8449b6337ef8b1957ea146da736011b67

Observation 9fb8845d-8e9c-4e18-981d-6214dbf83a0d · outbound

This paper cites Assessing diversity collapse in reasoning.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Assessing diversity collapse in reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.881955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.881955Z digest=sha256:a796d540c30da779a363feabcaa19931dfbe5412c6ba2403a91266f724d76cde

Observation fee6f11f-2095-4421-a497-0c9994c674eb · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Direct preference optimization: Your language model is secretly a reward model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.976302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.976302Z digest=sha256:20407bcedaca9675a0c5306aa0a8f6372f4706224e6b6aa8af36e3da06b7d480

Observation 699988d1-d4a5-403a-b834-fa3f35c6c9b9 · outbound

This paper cites Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:37.043476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:37.043476Z digest=sha256:a29c6d6fb0efc9d8548dfc84af6e4288ae87437219b714a606bf86513a824f01

Observation eaf1f6ed-4b5c-439f-ad65-cf6251c95323 · outbound

This paper cites Learning to Reason under Off-Policy Guidance.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Learning to Reason under Off-Policy Guidance

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:37.128523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:37.128523Z digest=sha256:3041a69e752191f62374d8b5f93aa5dd2b9bbf27b69e8b6df2e0bc5aae90d8c9

Observation 3294b235-12d3-4aea-af15-30508f14116f · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models HybridFlow: A Flexible and Efficient RLHF Framework

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:37.204240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:37.204240Z digest=sha256:926097cdc9a2b0cd4a11ae858907367f186e84f6f24aebe1cd9d3b32a47e105f

Observation 2cbc1157-af60-4496-8d47-c2d7ac995883 · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:41.227473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.313993Z digest=sha256:84c17ee88fe50793c98d0e209177c448d4c91b387d8ba1fad34675aff598c098

Observation 0b287539-eb83-4218-bd02-e499e193e097 · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:41.116242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.393926Z digest=sha256:5ec564eba6ebb97faae72385ecb1b7f640231a73addfda35919661b9a0bb11ec

Observation 3ed7525f-09d6-411a-9a85-7468a3918160 · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:40.987925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.478330Z digest=sha256:42b163695106b7532635151fd81df7412128e9a0b9312df6596189ac8d0feef0

Observation b6d942a2-c234-4217-98f2-f56e1763451c · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:40.908596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.558393Z digest=sha256:e8382cb31d65cba916d694b80d0798b811af460d2a5fbd2d5023ff7427b366b8

Observation 2bc9a944-19ff-4b71-bdeb-cc3b8a69008b · outbound

This paper cites A hint to the problem is provided below: [HINT_START].

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models A hint to the problem is provided below: [HINT_START]

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:40.776609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.607764Z digest=sha256:2eca558d5c2fe759e9462e1c354eb82ae1a733a2c867f601b3c3fe749fc285e0

Observation 85b570b4-9fb5-485c-b190-aab6de5316ce · outbound

This paper cites Think about how the identity (a²/4) - (4/a²) might be used as a building block for factoring the larger expression.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Think about how the identity (a²/4) - (4/a²) might be used as a building block for factoring the larger expression

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:40.646643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.657446Z digest=sha256:70893bc03af1dc97f1f92fa1b966b736f8359a431615303e6fcc385fc20da034

Observation e13d6c72-369a-437f-ad8c-5d16f7844303 · outbound

This paper cites Ask yourself if a difference of powers or a recognizable factorization pattern might help connect the two parts of the expression.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Ask yourself if a difference of powers or a recognizable factorization pattern might help connect the two parts of the expression

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:40.472946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.765087Z digest=sha256:7715dc37fda5e7b53ac4d6c4d017b5cf29bdd1eb994c5990f5b9efad11ca5a39

Observation a818a75b-c756-4365-9c30-8c403835893b · outbound

This paper cites How can this substitution simplify the structure of the problem? ,→ ,→.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models How can this substitution simplify the structure of the problem? ,→ ,→

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:40.307652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.850950Z digest=sha256:35046f92e48c39b8618a17a2d51dcdb017cc530e5f63f4ec29bcfd7fa401fde3

Observation ea984d5a-4322-41c4-93a6-29d8d3507596 · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:40.185924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:37.940379Z digest=sha256:072eeee551369343c3427b7b1ebfc3282867ec5046c2df65c409d4aebfdd3628

Observation 5ecf4a51-a6e2-4c78-aa2a-0f184f08c3db · outbound

This paper cites Use these observations to guide your step-by-step approach toward the final simplified result.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Use these observations to guide your step-by-step approach toward the final simplified result

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:40.020913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:38.063163Z digest=sha256:0339e3407de524ab080dc35e73676e013ad92d5b1d7fdf1f5b054b4665867128

Observation c4566f0d-4b2d-4f7c-a2b6-124fc1a5569b · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:39.842436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:38.128400Z digest=sha256:359ec7a61181db1406104585081b27374b06ac561a5202b42df3e232e1da0c8f

Observation d504db0d-c856-45e5-804c-5eb4622496fe · outbound

This paper cites an unresolved cited work.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:39.762269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:38.201603Z digest=sha256:5fe7642c17c109048f2f56e1f8f862fc3fff7ac12e4da2940ff329d915490f1d

Observation b03f974b-9501-4cfc-9c19-0e6a97c4b804 · outbound

This paper cites What does that imply about multiplying the side length? ,→ ,→.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models What does that imply about multiplying the side length? ,→ ,→

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:39.645444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:38.255892Z digest=sha256:d601ac4b45a80afcdfe42ed45730e139cc1c2e25063cb0b59951afca519aa245

Observation e8681102-13af-4343-9de8-61d3a7d873a2 · outbound

This paper cites Ensure each step follows from the properties of a square.,→.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models Ensure each step follows from the properties of a square.,→

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:39.495423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:38.345981Z digest=sha256:f498e9bd8d3726414c36bb4faf3da40b0203e4fe5d850054db4c0f34dd3630d5

Observation a5312c55-30df-4e3f-bd80-b29e1640dc0b · outbound

This paper cites using the hint.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models using the hint

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:39.367706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:38.429016Z digest=sha256:c4b5fd634ada3c3b67354e57f1cf614f0409012a60b995eb93ddb1a7d4e9a7bb

Observation 79e676a7-60b4-4be4-b507-194276321478 · outbound

This paper cites This example shows how a concise, domain-specific hint can redirect the model’s reasoning and correct a systematic geometric error.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models This example shows how a concise, domain-specific hint can redirect the model’s reasoning and correct a systematic geometric error

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:39.190927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:38.509431Z digest=sha256:2ac75d6ff4dfe55c7989e8bcff909c686502751ddca4814834c09728f2d24ad0

Observation 9efcd0cd-68c1-4bd7-bcb7-2855da000ad5 · outbound

This paper cites <think>\n {thoughts} </think>\n\.

Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models <think>\n {thoughts} </think>\n\

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:39.011456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T00:31:38.586086Z digest=sha256:7ad8792fbaf6488a1a117edf63e5c2ecb2a183fe9ca1992b1af8dcbc439fc9c3

Pith citing papers

Observation ae143971-845e-4b2c-a967-2259fae6503a · inbound

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning cites this paper.

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T23:56:55.008684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T23:56:47.274169Z digest=sha256:2f1d567c307e258d54c48fdfadb332131249be94f3ed9ec9c2accb3002f2dd63

Observation daff95ec-0fc9-4ba0-9be5-c2f08db63c37 · inbound

G$^2$RPO-A: Guided Group Relative Policy Optimization with Adaptive Guidance cites this paper.

G$^2$RPO-A: Guided Group Relative Policy Optimization with Adaptive Guidance Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T17:19:46.624230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:19:46.624230Z digest=sha256:2df627050ca0a823fba8291fc81198e1ed4b8afaddc7a262de6a31a0efdcff21

Observation a5a87b9b-3829-498b-a3b9-57b85d3d4566 · inbound

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards cites this paper.

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:26:28.367388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T14:24:48.666197Z digest=sha256:7523cc27ebc06ad51fc160fe84815857e74da474a394179ebf08279dc36260ca

Observation ab55c3cc-e7da-457b-9948-d3f526203de4 · inbound

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards cites this paper.

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T15:48:29.431431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:48:29.431431Z digest=sha256:262e20e403565f6efff50a5151021332c291b9b70f885dc1b5873fe447806f2f

Observation 11411b44-0b68-42f9-aeb6-eab51a369da2 · inbound

TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Relevance cites this paper.

TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Relevance Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T10:53:08.661292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:53:08.661292Z digest=sha256:cd8b37f4cadf3746af09141db49d2a3061e01aa42c0536a789fa5f46a6be314b

Observation c4ef3d49-52e8-4f36-943f-efa6b001fcff · inbound

Selective Off-Policy Reference Tuning with Plan Guidance cites this paper.

Selective Off-Policy Reference Tuning with Plan Guidance Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:47:05.199208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T01:28:18.615371Z digest=sha256:51d991ce93d447512584692979123e8d30e5f19312261e8455580a68993f58ba

Observation f6f35c7f-8dc4-4247-968c-0dd26bff82cf · inbound

Selective Off-Policy Reference Tuning with Plan Guidance cites this paper.

Selective Off-Policy Reference Tuning with Plan Guidance Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:22:59.342700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-14T21:20:24.066520Z digest=sha256:bbea28a550466502a459c3e0f9d25e552463e06fc6f57c509b17c660c367a9a0

Observation 2cf22fff-26b7-4055-817f-97f6907687bc · inbound

Learning Agentic Policy from Action Guidance cites this paper.

Learning Agentic Policy from Action Guidance Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:07:17.751601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T05:02:49.206053Z digest=sha256:9a18ac7a86b6eab99b5d5bacb48914ababb01373bcf11b5efb3660335bfd7d87

Observation 155a57b7-919a-4bda-89cb-84696c4310ee · inbound

Hide to Guide: Learning via Semantic Masking cites this paper.

Hide to Guide: Learning via Semantic Masking Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:34:39.107566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T12:16:12.108715Z digest=sha256:8a8358a2210b127fcb50ffb69312d241381853b36b5942a908c21807c876d6e8

Observation 9ca26ac3-6aa4-4d9d-99a9-9af34efbb577 · inbound

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients cites this paper.

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 94

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:48:56.211961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T01:08:52.981296Z digest=sha256:1023ade779a03afccd32622b003587060278207956568bb3d74a2ccae919ae34

Observation 8c9260d2-455f-405f-93e9-c5f57909abad · inbound

It Takes 8 Tokens: Weak-to-Strong Off-Policy RL via Auxiliary Branches cites this paper.

It Takes 8 Tokens: Weak-to-Strong Off-Policy RL via Auxiliary Branches Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T14:45:22.917496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T14:45:22.917496Z digest=sha256:ba13589f477eb8985bafcffcaaa4c03150c6886461fd838a724f64d9e37853f0

Observation 994c961a-99d7-4a9e-8378-21d7cd9ceb46 · inbound

LoRA Scaffolded Policy Optimization (LSPO): A Sampling-Time Low-Rank Scaffold for Recovering Reinforcement-Learning Gradient on Zero-Reward Cliff Prompts cites this paper.

LoRA Scaffolded Policy Optimization (LSPO): A Sampling-Time Low-Rank Scaffold for Recovering Reinforcement-Learning Gradient on Zero-Reward Cliff Prompts Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T01:28:56.693276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T01:28:56.693276Z digest=sha256:ea96ac336590ad3129e9945576f8db4c9f3480d8ed5fc9b6cfb01f96fa67770a

Observation 44f73572-76ad-4cac-a237-d43b714ec9a0 · inbound

Beacon: Knowing When and How to Perform Agentic Visual Reasoning cites this paper.

Beacon: Knowing When and How to Perform Agentic Visual Reasoning Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-31T02:45:28.621215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:45:28.621215Z digest=sha256:75860ec9094eceae9a0c6a0fe4f5525a322136c60f3a6024e2763c6e98b956a3