Pith. sign in

Paper Citation Record · LEDGER

When Does Muon Help Agentic Reinforcement Learning?

As of 10 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.16169.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.16169 v4

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T04:19:39.717440Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1c2244b-635d-41eb-bdd2-695eb4c5fb18 · outbound

This paper cites InProceedings of the 62nd Annual Meeting of the Association for Computational Linguistics.

When Does Muon Help Agentic Reinforcement Learning? InProceedings of the 62nd Annual Meeting of the Association for Computational Linguistics

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:37.337225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:37.337225Z digest=sha256:5764e3ed595b1e1240ef1163a2c7260a5cee6c5e164d9e178d493d88587e974e

Observation 9ab76eb5-1731-4888-82ef-7badbef85f51 · outbound

This paper cites ArXiv:2602.22817.

When Does Muon Help Agentic Reinforcement Learning? ArXiv:2602.22817

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:37.610753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:37.610753Z digest=sha256:c808729a529baeed6c8d2d327e75009bed68089e33d572cfbbb9303b46049063

Observation 480bcf6f-b36b-4529-b349-a8c5530ef28b · outbound

This paper cites MiMuon: Mixed Muon Optimizer with Improved Generalization for Large Models.

When Does Muon Help Agentic Reinforcement Learning? MiMuon: Mixed Muon Optimizer with Improved Generalization for Large Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:37.724140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:37.724140Z digest=sha256:3fc309879f6ba6af3bcff202b29b83e736432c37743f5a4ed2442a7bcbb0e3bd

Observation f0a3e926-f1ad-4016-8abd-f1b2aeac5ec5 · outbound

This paper cites SOAP, Muon, and Beyond: Pushing LLM Pretraining Scales.

When Does Muon Help Agentic Reinforcement Learning? SOAP, Muon, and Beyond: Pushing LLM Pretraining Scales

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:37.877565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:37.877565Z digest=sha256:9c435673073f1fedab9c636f987eb4085d4401fd952eaf5273e2d5e239f192c9

Observation 945767c7-a8f1-4104-888d-16bc9aaa08bf · outbound

This paper cites Lion, K.; Hübler, F.; Li, B.; Orvieto, A.; and He, N.

When Does Muon Help Agentic Reinforcement Learning? Lion, K.; Hübler, F.; Li, B.; Orvieto, A.; and He, N

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:38.180888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:38.180888Z digest=sha256:d194fd89dd74477dd01c1b59e053ddd94dbdc0d3e98b872f8fcbe631541c610c

Observation 5360b176-75a2-4f1a-99f7-d08c326dad6e · outbound

This paper cites Muown: Row-Norm Control for Muon Optimization.

When Does Muon Help Agentic Reinforcement Learning? Muown: Row-Norm Control for Muon Optimization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:38.316139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:38.316139Z digest=sha256:d6d844e18bd9ef1a58a38f808bd42f6d356b79dbb30add8dbabd547f05ed3dcd

Observation 31f5b1af-31b0-4aac-9ee3-6833d9bf061e · outbound

This paper cites Optimizer-Model Consistency: Full Finetuning with the Same Optimizer as Pretraining Forgets Less.

When Does Muon Help Agentic Reinforcement Learning? Optimizer-Model Consistency: Full Finetuning with the Same Optimizer as Pretraining Forgets Less

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:38.424519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:38.424519Z digest=sha256:f89d5b76f05f6bfc6996152fcf0ac0ac4d5da31b651221a1c7d1bdac6c8a45d2

Observation 53346f2d-09e9-4009-bfcd-963fb0f8f16b · outbound

This paper cites Muon$^2$: Boosting Muon via Adaptive Second-Moment Preconditioning.

When Does Muon Help Agentic Reinforcement Learning? Muon$^2$: Boosting Muon via Adaptive Second-Moment Preconditioning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:38.514523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:38.514523Z digest=sha256:e01c489b5825769e36fc5d8670121322d711f1bd27d0e56ce6a56dd91a8b1040

Observation d7029920-6001-4cca-a5b5-91c07e9f138a · outbound

This paper cites Meng, Z.; and Chen, K.

When Does Muon Help Agentic Reinforcement Learning? Meng, Z.; and Chen, K

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:38.715893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:38.715893Z digest=sha256:a36d184c28242ed4d0f87bbdcfbd6577a34305d9b739d31f600dd9cfc11882f4

Observation ddda07d6-d500-4e11-9822-d809b12448e7 · outbound

This paper cites CRAFT: Counterfactual Credit Assignment from Free Sibling Rollouts for Self-Distilled Agentic Reinforcement Learning.

When Does Muon Help Agentic Reinforcement Learning? CRAFT: Counterfactual Credit Assignment from Free Sibling Rollouts for Self-Distilled Agentic Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:38.920308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:38.920308Z digest=sha256:ffdf872d1657800bb3f44f15576179c441c53a5a3dd04bf21cd082ff6d8e1a90

Observation 9a2e4e88-102b-4c41-97e4-78687ea8ba3b · outbound

This paper cites HTMuon: Improving Muon via Heavy-Tailed Spectral Correction.

When Does Muon Help Agentic Reinforcement Learning? HTMuon: Improving Muon via Heavy-Tailed Spectral Correction

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.069342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.069342Z digest=sha256:c41c5ea4c0948d30a25132e53ac75afbbd6f4a8798ca47a0e49459cd8ae33b41

Observation b1eade41-a668-4039-9b4c-41f2e10aca10 · outbound

This paper cites Qu, X.; Huang, P.; and Horvath, S.

When Does Muon Help Agentic Reinforcement Learning? Qu, X.; Huang, P.; and Horvath, S

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.222709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.222709Z digest=sha256:e4dcb740085f45f495509460b55adbab6eaebb14abc00753fd0c6ade3ca12343

Observation 734d5f15-22eb-4cf3-ab53-819cfdfae2a3 · outbound

This paper cites Can Muon Fine-tune Adam-Pretrained Models?.

When Does Muon Help Agentic Reinforcement Learning? Can Muon Fine-tune Adam-Pretrained Models?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.354599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.354599Z digest=sha256:acbdac6812185572d2518fbd272f689150c89f984132f685ddc13a7f4d6a1493

Observation cffaf9cb-68a8-49bc-8848-b2a9d3007c82 · outbound

This paper cites EnvRL: Learn from Environment Dynamics in Agentic Reinforcement Learning.

When Does Muon Help Agentic Reinforcement Learning? EnvRL: Learn from Environment Dynamics in Agentic Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.533735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.533735Z digest=sha256:52655fb9a107b53eefecc8c28c7c42b5d706f7dd2f6f7ce2c1b4216249e602f9

Observation 7240a8e5-bb55-401b-8e5d-251fa0ed7eac · outbound

This paper cites Semantic Consistency Policy Optimization for Reinforcement Learning of LLM Agents.

When Does Muon Help Agentic Reinforcement Learning? Semantic Consistency Policy Optimization for Reinforcement Learning of LLM Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.560422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.560422Z digest=sha256:3b3729ada045bb47f0b76e1ed4391a0ecc5af0d783548673174eb2a1d9e5fc4d

Observation 9c92c0b1-3975-449c-ad1e-73d115cea13c · outbound

This paper cites StraTA: Incentivizing Agentic Reinforcement Learning with Strategic Trajectory Abstraction.

When Does Muon Help Agentic Reinforcement Learning? StraTA: Incentivizing Agentic Reinforcement Learning with Strategic Trajectory Abstraction

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.591203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.591203Z digest=sha256:1172e6f821c6ac0108fa9c436c231e65a0db3a024c7c68dee4fc7a012e8b189a

Observation 1da60b44-fa0d-46f8-af08-cc2b65a7413e · outbound

This paper cites Qwen2.5 Technical Report.

When Does Muon Help Agentic Reinforcement Learning? Qwen2.5 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.628532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.628532Z digest=sha256:411a2835730f145c196c5e699f408c1e5b2df7149933ddc18c45fc4ad7cbcd56

Observation a8fccb79-8244-48ba-b63a-72a5f0c9bda9 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

When Does Muon Help Agentic Reinforcement Learning? DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.663587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.663587Z digest=sha256:2a41f0ddadd637782dbfddd85f025d8a4a50bdb01187bd877162df93e2a1e0ef

Observation e22bd75c-120a-4213-9c03-f5f540262eab · outbound

This paper cites AMO: Adaptive Muon Orthogonalization.

When Does Muon Help Agentic Reinforcement Learning? AMO: Adaptive Muon Orthogonalization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.717440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.717440Z digest=sha256:4b0460fc0dd1320107da64c9188f837a3fb1c7586b17ab121a7cdeadef328fa7

Observation 9f85e988-4666-48a2-8404-79b41df5d528 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

When Does Muon Help Agentic Reinforcement Learning? DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:39.483270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:39.483270Z digest=sha256:66f28d61b07f79de6f6d5d3b65d01b4f8a6b0d7f2fde67930ef66a63e4e8e9b4

Observation 19c581b5-bdb8-4dea-8192-c779640ef290 · outbound

This paper cites Kimi K2: Open Agentic Intelligence.

When Does Muon Help Agentic Reinforcement Learning? Kimi K2: Open Agentic Intelligence

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:38.028406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:38.028406Z digest=sha256:a7128e0193cd562a4a94b3a92e0804dcaeb58ff1a484f0d90fd27e674e5d2e79

Observation bbe6a394-062b-4864-a496-6a5b3836f254 · outbound

This paper cites Rethinking Muon Beyond Pretraining: Spectral Failures and High-Pass Remedies for VLA and RLVR.

When Does Muon Help Agentic Reinforcement Learning? Rethinking Muon Beyond Pretraining: Spectral Failures and High-Pass Remedies for VLA and RLVR

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-04T04:19:37.465048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:19:37.465048Z digest=sha256:5a2584539fee20a233ffbd3008a85f9cdd698b1f2c6bd8b41eab1a34c94f6d3d

Pith citing papers

No inbound Pith citation observations are available.