Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:49:22.378838Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2510.24515.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:49:22.378838Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
18 of 18 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 66b07687-4b4c-42fa-a703-60781bbf25e7 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games The team orienteering problem,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3e00d25a-5dcd-4dc9-a7f2-51fe14d2c274 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Multi- robot scheduling for environmental monitoring as a team orienteering problem,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 173b04ed-e48b-4dbd-bb33-c23bfb0dbbc4 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games A spatio-temporal representation for the orienteering problem with time-varying profits,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4543c375-f62d-4c83-8294-82afcf7d960b · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Orienteering Problem: A survey of recent variants, solution approaches and applications,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e3b76725-514f-4b4d-8751-4792665916e9 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Learn to solve the min-max multiple traveling salesmen problem with reinforcement learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a93edeb8-854f-43f1-a770-fc79719d11b9 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Competing for the most profitable tour: The orienteering interdiction game
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3bf35bf7-251b-472f-8e9f-813795920025 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games DIRECT: A scalable approach for route guidance in Selfish Orienteering Problems,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation be888651-ed16-4863-95d2-58cee8539732 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Prize collecting multiagent orienteer- ing: Price of anarchy bounds and solution methods,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ec52a542-0848-49d4-a561-abf4f513fee8 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Collaborative dynamic scheduling in a self-organizing manufacturing system using multi-agent reinforcement learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 21345f24-b39c-497d-b30f-a0524eaf49fc · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Optimizing task scheduling in human- robot collaboration with deep multi-agent reinforcement learning,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 15a68cac-e63c-4ace-96e9-16ea94ed60cd · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games A dynamic task assignment model for aviation emergency rescue based on multi- agent reinforcement learning,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 544bfbf2-e66d-4aa4-b7bd-78c90ca65377 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Graph Attention Multi-Agent Fleet Autonomy for Advanced Air Mobility
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fea77039-023c-4e29-9f7a-d85a3953bf64 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Consensus-based decentralized auctions for robust task allocation,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a6967ef-ab77-44bc-9b7b-31c331827c8d · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Optimal cost-sharing in general resource selection games,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b88d9933-c6fa-4966-988f-7a3b2d65354f · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Nash q-learning for general-sum stochastic games,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 543bf2e6-9c22-40a5-9b9d-132ef7e07659 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Multiagent Systems: Algorithmic, Game-Theoretic, and Logical Foundations,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cdeba35a-1efa-4ffe-afcd-fa9a50b79ee1 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Stabilizing Transformers for Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9332d9e6-77ba-4d08-9693-a3037e297754 · outbound
Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Orienteering problem: A survey of recent variants, solution approaches and applications,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.