Pith. sign in

Paper Citation Record · LEDGER

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2607.24833.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24833 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T07:34:57.251159Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved48
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 81b18e99-c7e4-4435-892c-76782d9ce4ee · outbound

This paper cites 2026 , note =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2026 , note =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:52.667421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:52.667421Z digest=sha256:442824cca8b3788ed9ab93424787ffe241894f675a6bc77b4bb680bd1e1436a0

Observation 244961c9-8f48-4f6b-8f4c-cce26612c28c · outbound

This paper cites 2025 , note =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2025 , note =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:52.720900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:52.720900Z digest=sha256:2600eaba28c0b601ad3591af33d94689308768ac1278087e114bd985c93563fe

Observation 04b29a83-3f63-4861-b4e5-29a1d2c9d395 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:52.808847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:52.808847Z digest=sha256:454916a764990493f15dade0b21c3ffa8301d1e7dc465f39afcb0b7a8c2ffdac

Observation ef391fe9-6136-4169-a24a-8f0556461d0b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:52.905598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:52.905598Z digest=sha256:2db45594f90f3f6d56b842d4a0c6753bf661cfea3b877dfa4a144d2f527ad3ee

Observation 6bacec8d-f120-4ce3-a2b7-db03af9e9a12 · outbound

This paper cites 2024 , note =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2024 , note =

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.022906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.022906Z digest=sha256:ecf9d46d5c3512124adfd7f6de80f5db8fda837372cea3de30cb621c81d14835

Observation 87a9d1a3-6d2d-4968-946a-4afa7233e0eb · outbound

This paper cites Proximal Policy Optimization Algorithms.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.089892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.089892Z digest=sha256:41ad3de472ceeee6e913d01d496b9a51ed6c719a630d35ca9bdf0edb9318ed36

Observation 4cd58a9d-cc36-4cf6-98aa-2e3635c3c67f · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.172408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.172408Z digest=sha256:5cdac20bdfd5dff604cb0f014c87cdade000ab01692fe9689fce8d6042f1c6c5

Observation 35971f69-dcd6-420b-b6df-2ec85b50d6cf · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.287453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.287453Z digest=sha256:9bb2c4cc6953db7d1312365055717e09fe9f4b36440a0c09874fafc9bda98b2e

Observation 56b98f94-39ec-4418-b207-d4eaaeecbfb5 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.352958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.352958Z digest=sha256:1eb92112282ef0f89078cd3543670c6e767cd92dbf3b8cf8bfc5b4099684bb86

Observation 03fce6f1-890c-438b-91ae-5704d09df196 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.462245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.462245Z digest=sha256:735971ba844bc380be04fa9c1204bdfcff17331f05642cd4c4e12051d023630f

Observation 6a91bba5-0c5d-40b4-823a-56c3226fb3c6 · outbound

This paper cites , booktitle =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning , booktitle =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.578840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.578840Z digest=sha256:f1d3754f95326fad7f513c06ec9661ad71d5456c645f4d1f32bb48739efe9d9a

Observation 501596e3-1294-47b5-ae47-ff65e4dff05b · outbound

This paper cites , journal =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning , journal =

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.652970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.652970Z digest=sha256:418311ec3002e9d686b3d13ffe8b71916d4f2b016cc3b2f5c29552dc1ea9a0de

Observation ebb1ed94-b1b7-42a9-8472-a027ac6eb1e8 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.721806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.721806Z digest=sha256:3f1350496461fb28508d3b5a19be9cc2c30bfe35ac43630ff9b1d90d14967a24

Observation caa7f80d-8fa4-40ef-9fb1-2d1aef61493f · outbound

This paper cites and Zhang, Hao and Stoica, Ion , booktitle =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning and Zhang, Hao and Stoica, Ion , booktitle =

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.842119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.842119Z digest=sha256:4853f30e3c318702a403692100dc61f90e240efe7470c9694a2439988c60d2d9

Observation 1df552fc-4245-4de6-ab8b-e5cb5d41a851 · outbound

This paper cites 2025 , note =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2025 , note =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:53.929243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:53.929243Z digest=sha256:09572e32ca6bafb7e7d658ab6b3cef44796554a89608453e1426927c8c1521e8

Observation ecadc99e-6b6c-40d1-b1cd-007ea25f2f97 · outbound

This paper cites 2023 , note =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2023 , note =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.016895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.016895Z digest=sha256:2f8ffc87b5a9cb35b4d05a4d6eb8f1e6142b3b05281c61e782df3daebc8f9308

Observation 27bc778d-a9ab-4877-a035-5bc786942a90 · outbound

This paper cites 2024 , note =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2024 , note =

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.072903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.072903Z digest=sha256:df9cd1175c5590b7f5d97aa059e44c95f0d42b6fe99b72b4cacd174fb0075447

Observation 5b691810-3838-4793-961b-6550e5a5eb06 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.177432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.177432Z digest=sha256:718ef79e7296ceb204db7fedc12851afdc56897a1deed3ece6ed1f9387e3a0b9

Observation e7a4c51a-2945-45bc-a7bd-2a6e11a2109a · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.255689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.255689Z digest=sha256:8407c7a7de4f2ced9d115acef22b6c2c8c5984b968761fdbd2f5a1e4b4563325

Observation 8569224e-5a1b-4503-bc37-8c8a8ad3c12c · outbound

This paper cites International Conference on Machine Learning (ICML) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Machine Learning (ICML) , year =

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.337048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.337048Z digest=sha256:44797e31ed1615aac45acb0609b43b5599e21f6c99db3b2a83aa9606d4a8b775

Observation a0fd8522-fee6-4c9a-b44e-29d9c7ae87e5 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.455969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.455969Z digest=sha256:dc9410430df03547db6f485bd8bbed8fc7341b38c38a87987e977ff9d233161c

Observation 7d5b5313-5282-441e-9b9e-0622dcf199e2 · outbound

This paper cites Don't Tell the Answer, Truly Guide the Reasoning During.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Don't Tell the Answer, Truly Guide the Reasoning During

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.537997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.537997Z digest=sha256:14e1a37d26420c7d208cb57ea6651ca8c4afe2badb39f323d605af6334bb6917

Observation 8c9d8cfe-5ecc-4540-9d30-8741ced11c91 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.601137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.601137Z digest=sha256:76984d1fb28d641903e0afba198c3f4260b55847c864d897d24c6169c3be4386

Observation 45429833-a9ff-4208-938f-93cba4343624 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.701839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.701839Z digest=sha256:6f2f968b124a20df7d6d61123df5821db5672b91cc963934d18d57a0082350c6

Observation 26c06b32-36c7-43d3-ba09-8a252be4e03a · outbound

This paper cites Staying in the Sweet Spot: Responsive Reasoning Evolution via Capability-Adaptive Hint Scaffolding.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Staying in the Sweet Spot: Responsive Reasoning Evolution via Capability-Adaptive Hint Scaffolding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.765756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.765756Z digest=sha256:a1b0a1c4f7c46d7f6e4360eeb7216764ecda476985842f6943cde07ed1a8623b

Observation b6635fc8-6287-485f-9f4f-f0b72cdaf244 · outbound

This paper cites Boosting.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Boosting

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:54.942938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:54.942938Z digest=sha256:be78dd5adeba8fb351ebda527e164f0eef6845bc9f72942fd3cc0e26863045b9

Observation 3dfea157-c72a-4a0e-aaac-d803c40e3597 · outbound

This paper cites Train at Moving Edge: Online-Verified Prompt Selection for Efficient.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Train at Moving Edge: Online-Verified Prompt Selection for Efficient

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.055430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.055430Z digest=sha256:014e0a2952e191cbe71b857f7d077544879a8f87952deb6ae961b4d3a39eed19

Observation 19fca2ce-0a71-4027-b71a-d83b82273fa7 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.112071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.112071Z digest=sha256:bd878b9ec42e8f51e744ec5665d145800240a306ddf917328ac6f4c2e4be72b6

Observation 0733587a-51d5-43fe-bd4f-8fb3e44c8c28 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.220532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.220532Z digest=sha256:71aade30e0b7fe451490b78bda5b488bd54fddb6d123cd43bfad730bece1037f

Observation 6156148b-691e-451a-9ee1-c8ad2f6d81d7 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.301544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.301544Z digest=sha256:6c1ae2f82666f85e107f902d830eba7dbd1aac94995b4bcb7edb90bad1590952

Observation c30af885-3f78-4ee6-8acf-9c45f11113fd · outbound

This paper cites Measuring Mathematical Problem Solving With the.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Measuring Mathematical Problem Solving With the

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.388018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.388018Z digest=sha256:9ffddf709b8bbcc428776968d198fe4792631f78ebb1bcc6550afd510ce292bb

Observation 2172d4fc-2eaa-4f1a-bd87-a8cd7a30081e · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Training Verifiers to Solve Math Word Problems

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.532748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.532748Z digest=sha256:1d01836eed6c9b134433e805c8ea2023540fd1ce560100e3ba8f18e4ee16ec27

Observation 30538229-4d0d-43fc-8c6b-b9800ed7363b · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.602336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.602336Z digest=sha256:c5c573abe4e6e4f5ddbad7f3aabe92401ee1f229c59b3a0a1e820b80c182e9aa

Observation 02d054e1-546e-4910-b105-ac5b7fbd7c13 · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.683543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.683543Z digest=sha256:ee597566cc2846e6ce87defbb856cc16713cc9bcbca167fa019516e39f4c9c8e

Observation 036d2c91-1866-401b-bfe9-3eb6cbb782ba · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.763796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.763796Z digest=sha256:26eee0bc34624bc1294163696d2e0f2b1579726d4c70dbf1ec1e60837d3e7bd5

Observation d372dc61-e76c-4364-9a82-5502f9bf78b9 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.870328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.870328Z digest=sha256:4a7aaa523dcad4d7d11ae9892b94eae2aa357bce2b7d89673ad398d793922e4e

Observation 72e1f0de-0a34-46c2-a54b-8f58af08b9a5 · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:55.960507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:55.960507Z digest=sha256:54d84b864a35d7e1ae7cb9cf4cba2fd270198d975ad2f6341d13b8329e0ac22b

Observation 0b6b5067-8c38-4037-b5e7-595390f58244 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:56.046520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:56.046520Z digest=sha256:e8625bf074310d1d52bd6f2dda717a84df08149a24e487096f9993b32d469785

Observation c956bc34-fda9-4db0-b9a6-724bbd67aa86 · outbound

This paper cites Back to Basics: Revisiting.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Back to Basics: Revisiting

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:56.122271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:56.122271Z digest=sha256:ae12ae51f3b5bf888a2a929d8bc2d1147905b2970fffb47dd716c79d0079bb1d

Observation 53302985-149c-4556-896b-0afd4b18c722 · outbound

This paper cites 2024 , note =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2024 , note =

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:56.231813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:56.231813Z digest=sha256:f97140ba2ee24eafa806e9e13824746d44eacef739e367fe51179901d25d8570

Observation 897bbd8e-ef9a-4d1b-a8bc-958357b59bb2 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:56.376887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:56.376887Z digest=sha256:8c6ad551d82562d9166eec398ff39486aebaabd830d96e9fd73b5f20d56d1b81

Observation 1535bde9-33fa-4139-9187-8d966b499f80 · outbound

This paper cites Reinforced Self-Training (.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Reinforced Self-Training (

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:56.525374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:56.525374Z digest=sha256:8a2b06f0a0dd047010206c18c8e691533779d1cb454be68ae86ae8edc2ef334a

Observation 2a60d578-c1c2-4792-8300-0387304d60ce · outbound

This paper cites Transactions on Machine Learning Research (TMLR) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Transactions on Machine Learning Research (TMLR) , year =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:56.646520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:56.646520Z digest=sha256:fbfd55b04c687f0041261fc1c8d855efb7d37e5292bff094aedbb61fd0c319c1

Observation 46f0c937-c1c1-42b7-b42e-7b9169cd431c · outbound

This paper cites Teaching Large Language Models to Reason with Reinforcement Learning.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Teaching Large Language Models to Reason with Reinforcement Learning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:56.790463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:56.790463Z digest=sha256:aab3739badffb2eedbd02e135965569012d31b563475e29fb426edbdad6d2b2a

Observation 389b8b2e-45e0-4354-bbc4-4edca3600dab · outbound

This paper cites 2016 , note =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2016 , note =

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:56.971940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:56.971940Z digest=sha256:fc2a3426666ef35a279fffebcd3bb7ab8cd696f49bef9b58559ace8501068ffc

Observation b79782fc-1867-4cda-847b-521e8cb4e170 · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:57.087681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:57.087681Z digest=sha256:758a110ee83b2fe849f1abd92aa9b6788a2bd4e1fd45c41d403718d435eaf310

Observation 294c8b0e-9dfd-4cb3-8a34-8b5fa2165ff2 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Evaluating Large Language Models Trained on Code

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:57.162319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:57.162319Z digest=sha256:257b7cafe0980b985c8ddea3c5aa62a5e87a7337e472a6e228950b22e1ce2ff5

Observation f533e7ed-2f38-4c6f-8b49-9056c522fab1 · outbound

This paper cites an unresolved cited work.

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T07:34:57.251159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:34:57.251159Z digest=sha256:55e95922c5a959e6dd92d6b83bea9bb3e88e86c5b92a6a82d3cd25885ec10906

Pith citing papers

No inbound Pith citation observations are available.