Pith. sign in

Paper Citation Record · LEDGER

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution

As of 9 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2607.26784.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.26784 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T21:12:59.453856Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T19:53:25.916475Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T19:53:27.058819Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c1fe03a0-f688-48a9-bdc8-d2d1679c3821 · outbound

This paper cites Group-in-Group Policy Optimization for LLM Agent Training.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Group-in-Group Policy Optimization for LLM Agent Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.359390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.359390Z digest=sha256:6c5b2948d0887f95662251d4357a7078aa19659178873bb735b7d3de7002f3e1

Observation c9c878b0-f740-42ec-bfb0-593fc9ed2b90 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.366710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.366710Z digest=sha256:8385743366a6febc9016132dc69dda1cb7f3d9ad4b9fb642d91741b7e656acd8

Observation 5fd41044-b57f-451d-942c-c34a313e79b7 · outbound

This paper cites Hierarchy-of-groups policy opti- mization for long-horizon agentic tasks.arXiv preprint arXiv:2602.22817,.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Hierarchy-of-groups policy opti- mization for long-horizon agentic tasks.arXiv preprint arXiv:2602.22817,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.370562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.370562Z digest=sha256:ec70d79c7f34c5b9bf8444e01f03bbdaf03f9ead7740b31b06f48066adf6433c

Observation 2f4c6aba-b412-47bf-820e-799e47818d37 · outbound

This paper cites Meta-rl induces explo- ration in language agents.arXiv preprint arXiv:2512.16848,.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Meta-rl induces explo- ration in language agents.arXiv preprint arXiv:2512.16848,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.373942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.373942Z digest=sha256:670461f70034efea027654b6f31824d421533d935772507c0cf4ff4cfad11ca5

Observation a7576976-63a4-4c83-89a4-269ee79ae7c9 · outbound

This paper cites Self-Distilled Agentic Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Self-Distilled Agentic Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.377118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.377118Z digest=sha256:e83ea57df8d06edec7a47d7df4d2cb3bfb8ba09061a02ef311a0f8eda129095e

Observation 52253690-8d56-4915-ba29-1014e81f4c93 · outbound

This paper cites Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.381111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.381111Z digest=sha256:ff09c0424fce27432e48ab49350c1aa876b1d081658d5cf4196ef374c3288491

Observation 91a61321-88af-4969-ab41-b190bc5386db · outbound

This paper cites ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.384535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.384535Z digest=sha256:d727ce445fbd3b2af48910fb5a5598c99fa0f16009935a0454d21610e49d3359

Observation 76ddb756-05b0-4672-a88d-6b70121e391c · outbound

This paper cites SkillOS: Learning Skill Curation for Self-Evolving Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillOS: Learning Skill Curation for Self-Evolving Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.388053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.388053Z digest=sha256:041768af03f1ccdf06726d162510299106b0776f41e0a9dd07cbc0778092dbf2

Observation 621de15e-e44f-4f77-91ff-bdeb1b8acde4 · outbound

This paper cites Webrl: Training llm web agents via self-evolving online curriculum reinforcement learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Webrl: Training llm web agents via self-evolving online curriculum reinforcement learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.391280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.391280Z digest=sha256:03e71ff4bfc44d9abd5913bdb0c221b7d3bf4dbf540d3bc53d786fd844f00495

Observation b14af90f-e9c6-4e11-bc14-f2653b3e5646 · outbound

This paper cites Autorefine: From trajectories to reusable expertise for continual llm agent refinement.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Autorefine: From trajectories to reusable expertise for continual llm agent refinement

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.394453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.394453Z digest=sha256:acb561d42ff60fda5e28a26a3182779292d2842773325fa26e0f035ffca7ad10

Observation 051619d9-698c-41d2-a156-23f3a22854bc · outbound

This paper cites Proximal Policy Optimization Algorithms.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Proximal Policy Optimization Algorithms

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.397336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.397336Z digest=sha256:944ceab47cb421f10f4bf786443296153d6b2f1cec7fcae63ead75e59a07af0d

Observation 06dce564-2ccb-414d-864c-37813649e9cd · outbound

This paper cites Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.403621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.403621Z digest=sha256:556ae728c29f181ff18d04fdf421d5136040bcf2153f62babd8600cb819d206b

Observation d161b1b6-f3e8-44f0-8372-1895d21d9b32 · outbound

This paper cites Milestone-Guided Policy Learning for Long-Horizon Language Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Milestone-Guided Policy Learning for Long-Horizon Language Agents

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.416855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.416855Z digest=sha256:0b96b9182973d07f33299bac49d912d1c55c9e472308e128a99f0366892f1707

Observation 2dc2ff19-2d82-4a25-b66a-44a6ea5b63dd · outbound

This paper cites SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.425331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.425331Z digest=sha256:5c5d8687e69c3010be32ceade423c61c347f05b4cb3e71fc68e78db1d3956b39

Observation 4e2c48f9-e951-48b3-8d4a-71238d638059 · outbound

This paper cites AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.428554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.428554Z digest=sha256:07e3c24107a0f9fed94e5ecf1e3a05b118998b85aab304219fddf97924dfae2b

Observation f0700bb3-ee9b-499a-80ae-c4b63f8744ee · outbound

This paper cites SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.431570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.431570Z digest=sha256:77232760e9cdf14ffd882400ac58ec696b3b0126d13edf744f94c1b111517c6a

Observation eac22b0e-9bf5-4cfc-be3a-592faf33dd12 · outbound

This paper cites Qwen3 Technical Report.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Qwen3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.434695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.434695Z digest=sha256:6a8d9b458a42d7e7c2aab5870a1d2ad4e458a368a6e89be31f50b6074ee8f366

Observation b61be434-2d06-46dd-aa2b-719bbe066079 · outbound

This paper cites SkillOpt: Executive Strategy for Self-Evolving Agent Skills.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillOpt: Executive Strategy for Self-Evolving Agent Skills

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.437826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.437826Z digest=sha256:b0da79d5cb8d8f05d00eb65ae135bdeff61da86ee590b8c2633952469c18f008

Observation 5eaf3873-1ce2-4aef-ac20-0b124d42ce0b · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ReAct: Synergizing Reasoning and Acting in Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.440916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.440916Z digest=sha256:b5f4dcb388737e83b32cd7955c0a9da0f8a6dad22dfa72d02937faf8d24383f8

Observation b2e17167-aedf-42f2-af8b-a9e7e015f472 · outbound

This paper cites Look Before You Leap: Autonomous Exploration for LLM Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Look Before You Leap: Autonomous Exploration for LLM Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.444076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.444076Z digest=sha256:4110ee2f4061a016a54736a214855db6399c617ff50e38e9f332db94b0361cbd

Observation 7a145b9c-9181-430b-acfa-2b13de815d79 · outbound

This paper cites The Landscape of Agentic Reinforcement Learning for LLMs: A Survey.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution The Landscape of Agentic Reinforcement Learning for LLMs: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.446937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.446937Z digest=sha256:ee029a730fa22c0a4485d01f32957c33604bed12a7d1db210da425bef2ff20f8

Observation c31621b1-d508-40a7-9c59-a565a6f3e871 · outbound

This paper cites MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.449832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.449832Z digest=sha256:4e5e7406db0c8e1be46f73165f566c1e8ade360076b5472280a7b2a5b39d540d

Observation 884533dd-6334-41ee-af77-a2f36bc6e103 · outbound

This paper cites LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.453856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.453856Z digest=sha256:12625392196d8413e258f0ae253fe1693a4a6ca7c559f62024489c844664578d

Observation ce5cf7ee-9bd5-4544-bea0-a63c0043f712 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.400414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.400414Z digest=sha256:983c9c5a547b4df72ab4f7fb0de6c0f2c4b051e324fb7f63319bcb1ef440745a

Observation 9cb6b808-b81d-4065-a5ec-48e0eaf66b4c · outbound

This paper cites Reinforcement learning for self-improving agent with skill library.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Reinforcement learning for self-improving agent with skill library

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.410183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.410183Z digest=sha256:0ebc6dbcb58ef7c86eaca18af1d030a6926c85edee9adce29f07034118cbdea9

Observation 4af1e024-fd0b-47a0-a885-242921623c03 · outbound

This paper cites RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.413803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.413803Z digest=sha256:57e1ae063f3374a9e1dcebaa5119ec18f9f5746878a40c2b14e4fed90337d4dd

Observation f822171c-725d-4a6a-afe9-60b799392731 · outbound

This paper cites ALFWorld: Aligning Text and Embodied Environments for Interactive Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ALFWorld: Aligning Text and Embodied Environments for Interactive Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.406582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.406582Z digest=sha256:1edc8b3bf22d06fddb7c2f9ccb40dd46367c67d8aa0caa9d75212e0c0e36c9a5

Observation 7a8d2d8c-6f80-4231-979d-dcf2680be10e · outbound

This paper cites A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.362919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.362919Z digest=sha256:cfb42b43bbd65eba1368b29bd1765de3e766fb7623355c032b2bf072abaae8d8

Observation 0701f85b-1ea8-4b58-99bc-4f33f420f26e · outbound

This paper cites Yuxin Chen, Yu Wang, Yi Zhang, Ziang Ye, Zhengzhou Cai, Yaorui Shi, Qi Gu, Hui Su, Xunliang Cai, Xiang Wang, et al.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Yuxin Chen, Yu Wang, Yi Zhang, Ziang Ye, Zhengzhou Cai, Yaorui Shi, Qi Gu, Hui Su, Xunliang Cai, Xiang Wang, et al

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.351192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.351192Z digest=sha256:5d2e745d305e80e4b67bc7998a7a92489f90f257e415382d3922157d22456a1e

Observation 4d9ba3fa-3322-402a-a4ca-e715ccc16e2f · outbound

This paper cites Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.355533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.355533Z digest=sha256:1665e3eb1a1cb9dcd0f0a1834f504966a840d8b4aafa8ee5dcea144246cceebd

Pith citing papers

Observation 711881f7-a651-4888-8f97-894bf50f5a8a · inbound

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning cites this paper.

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution

Reference 66

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T19:53:27.063688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T19:53:25.916475Z digest=sha256:247398a4e4ccee1381f3ed048b299a7873aa9f90354153e91068f209dbb7e97b