Pith. sign in

Paper Citation Record · LEDGER

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning

As of 12 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2502.05537.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.05537 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:03:05.612017Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T04:53:10.425398Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T04:55:54.390079Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact1
  • verified fuzzy44
  • unresolved12
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6c77021-e3ee-4112-b584-715da3499ddf · outbound

This paper cites Hindsight experience replay.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Hindsight experience replay

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.542013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.324922Z digest=sha256:bb02d79d9675958b0f7d25a3729d771b1a99b916f4bd31459ac60d92a326bbde

Observation 80bc2e2b-6d14-497d-bdfc-a089e43c1f27 · outbound

This paper cites Modern graph neural networks do worse than classical greedy algorithms in solving combinatorial optimization problems like maximum independent set.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Modern graph neural networks do worse than classical greedy algorithms in solving combinatorial optimization problems like maximum independent set

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.329854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.329854Z digest=sha256:f56782240bdd1cbb39e8d7e30e4fe5027b59e8a4c6ace214b61744832b9a9d4f

Observation a3983750-a7ef-4efb-baee-21a3f9c5dd32 · outbound

This paper cites The option-critic architecture.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning The option-critic architecture

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.515774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.335054Z digest=sha256:ca9e4db46b218c687d38db4622252dda77985cafaa3c2cc3e9288da018e16c9a

Observation 1a98aa2a-c9a5-4ac4-81ab-4cc5cad97071 · outbound

This paper cites Neural Combinatorial Optimization with Reinforcement Learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Neural Combinatorial Optimization with Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.340426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.340426Z digest=sha256:905d06e99f301bbde052640cfe51fb4da3e45d3ce7fa8897386e25036be9a6d4

Observation 346bc664-227d-4ce3-a8de-a8a07546b3b2 · outbound

This paper cites Machine learning for combinatorial optimization: A methodological tour d’horizon.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Machine learning for combinatorial optimization: A methodological tour d’horizon

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.499738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.345801Z digest=sha256:555206bd1348753ad5d5c18824b026e3ece2482abd7ce80840791b21c4585fe6

Observation fc86a519-a3bf-44ae-8b0b-2c73b28e54d1 · outbound

This paper cites RL4CO: an Extensive Reinforcement Learning for Combinatorial Optimization Benchmark.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning RL4CO: an Extensive Reinforcement Learning for Combinatorial Optimization Benchmark

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.351089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.351089Z digest=sha256:dc6573696e326887d8e74861d0aa4d6b4351d75dad3a32c1b954b3dcff6d29b4

Observation fb027c3d-feae-4b3b-9fa6-fc9845017877 · outbound

This paper cites Models and algorithms for combinatorial optimization problems arising in railway applications.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Models and algorithms for combinatorial optimization problems arising in railway applications

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.483727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.356856Z digest=sha256:a51a849cc1c46409059a417e29407d183880fac177e3a8edb9e469bd5ce363ee

Observation ee40282e-4239-4fe5-9e56-a68597b703dc · outbound

This paper cites Applying gis and combinatorial optimization to fiber deployment plans.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Applying gis and combinatorial optimization to fiber deployment plans

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.468790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.361449Z digest=sha256:fee5b45986a5328f8c7b3b340f8ec1b20016184a95613efc8f66df726a2c02f7

Observation 521f4a51-8597-44ec-81a6-908180e21947 · outbound

This paper cites Improving optimization bounds using machine learning: Decision diagrams meet deep reinforcement learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Improving optimization bounds using machine learning: Decision diagrams meet deep reinforcement learning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.454089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.366486Z digest=sha256:2cfa5024ac8610ffa52361bb61fe46501985b909ae9672f63ebea23b8e844f20

Observation 3954d364-95aa-4a94-964d-af507c0c7b21 · outbound

This paper cites Contingency-aware influence maximization: A reinforcement learning approach.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Contingency-aware influence maximization: A reinforcement learning approach

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.439691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.371644Z digest=sha256:aca7cd08e3aa5b38b2ffb746f3bb4968865754e4c985c38b1245612e7720fe5c

Observation 6be00938-74af-433d-9b9f-c98995bc0d09 · outbound

This paper cites Learning to perform local rewriting for combinatorial optimization.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Learning to perform local rewriting for combinatorial optimization

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.424938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.376450Z digest=sha256:8a1eab8331ecb9b25357bfc06a84c4875dd131a2fba81849a8935bfc854f0e8f

Observation 405cfe16-e1c0-4ae7-a095-5ad040e51962 · outbound

This paper cites Discriminative embeddings of latent variable models for structured data.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Discriminative embeddings of latent variable models for structured data

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.409130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.381049Z digest=sha256:feeddb7bc913aaf3413675b19d6e3dd999b5893b56eca241f4885dd0d8a91b3a

Observation 41e48dd5-391d-42c7-af87-dc38750570eb · outbound

This paper cites Feudal reinforcement learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Feudal reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.394011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.385410Z digest=sha256:bcd6e90a994c2b2d4df455c4756031ed013184af49ae8a03242ff239a7848574

Observation 44c5ad8f-c132-4ef7-a2f7-b547bc87892f · outbound

This paper cites Learning heuristics for the tsp by policy gradient.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Learning heuristics for the tsp by policy gradient

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.377570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.389827Z digest=sha256:81206c3fe5271a9d1dba4956ad00d0a5aa56318bd0d3de6da4426643b09276e3

Observation 6fe79b74-02fa-480e-89a8-c4f60958210f · outbound

This paper cites Learning Permutations with Sinkhorn Policy Gradient.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Learning Permutations with Sinkhorn Policy Gradient

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.394095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.394095Z digest=sha256:eb6b8e02236dd406ca9e7cec552d96d55183d2c4eb9aeebab28634ddcdfefc4e

Observation 74691e8f-3ad9-44be-a04c-ad897a778d18 · outbound

This paper cites Generalize a small pre-trained model to arbitrarily large tsp instances.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Generalize a small pre-trained model to arbitrarily large tsp instances

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.361992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.399133Z digest=sha256:e4f9ca58d1ed6d25cf120faa692c1a5fdf5c1cea6356152444f966dd1b39bff3

Observation aa9ad8b4-5f2d-45dc-88e7-2bfdba7dc047 · outbound

This paper cites Deep Sensitivity Analysis for Objective-Oriented Combinatorial Optimization.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Deep Sensitivity Analysis for Objective-Oriented Combinatorial Optimization

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-08T19:03:05.709732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.403352Z digest=sha256:d2d050a37f0532040f34fc9b2e00352f37981d50e13c1c6b2d58adca1f762982

Observation bbf9f8ec-f50d-4c7f-94d2-e164ca2ec384 · outbound

This paper cites node2vec: Scalable feature learning for networks.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning node2vec: Scalable feature learning for networks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.346542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.407984Z digest=sha256:688c7e7501480f4e6f87a281f89ac84b4930f5b4fe3a81d81923c8d7aab32d95

Observation 92a868bb-bc63-406c-80e1-5ec618317c80 · outbound

This paper cites Double q-learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Double q-learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.332009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.412203Z digest=sha256:a3258fcb1fb1dbe1bb73dae15d91899452549cda0475a364ed05c81f67b88c1a

Observation e64f3668-ab28-4724-a84d-f348a2ae5b93 · outbound

This paper cites Efficient active search for combinatorial optimization problems.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Efficient active search for combinatorial optimization problems

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.317716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.417071Z digest=sha256:c4ac37cdc2a5d889badff0cf927af95d5795fbf4f0fa78aa93fd1ea25f297326

Observation 40361bc2-72a8-449d-8805-1839b7c52e35 · outbound

This paper cites Convergence of stochastic iterative dynamic programming algorithms.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Convergence of stochastic iterative dynamic programming algorithms

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.302762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.422052Z digest=sha256:3f1c194971e1ac2ee6a5105eea9b3984dc3886cda61de2727b39c73ba6c2ef6f

Observation 0a24db4d-80a0-4aab-965b-bac1f1209523 · outbound

This paper cites Deep reinforcement learning approach to solve dynamic vehi- cle routing problem with stochastic customers.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Deep reinforcement learning approach to solve dynamic vehi- cle routing problem with stochastic customers

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.285983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.427074Z digest=sha256:b1d2e4621afe853a9229c828082e1c466d2ae16361e5f3e1613aaae0c4af30f1

Observation 9a21f23e-ea64-4fd5-b9ad-2ed62956359d · outbound

This paper cites Maximizing the spread of influence through a social network.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Maximizing the spread of influence through a social network

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.268148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.432052Z digest=sha256:d47827a39e5f6e43452da90a4079c5a14a89da76c34121d2c48d6802e009495a

Observation 2785ac3a-62da-4da8-b57b-e8fa6f42d1f1 · outbound

This paper cites Learning combinatorial optimization algorithms over graphs.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Learning combinatorial optimization algorithms over graphs

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.252054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.436959Z digest=sha256:3161e88212444bd13e89908d131771f4593766b1e70a2b0dc03740a1f7ed2cc5

Observation a2bb0535-2d7a-4016-9f72-c87ce1aecf1d · outbound

This paper cites Semi-supervised classification with graph convolutional networks.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Semi-supervised classification with graph convolutional networks

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.235822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.441860Z digest=sha256:88b317ace482e83812ac9de7c6a83766ebee96ee69e8f5320fa2cc49869a8ead

Observation 1b5389a6-f425-482d-9f06-05742a76a4ca · outbound

This paper cites Attention, learn to solve routing problems! In International Conference on Learning Representations, 2018.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Attention, learn to solve routing problems! In International Conference on Learning Representations, 2018

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.219988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.447001Z digest=sha256:a527a5d6343a978bf2fe36f7d19477baf44a04c17b805ea44ff85932a20544c6

Observation 4d7c55bb-7df6-4aab-82dc-e25f51a54107 · outbound

This paper cites Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.203778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.451936Z digest=sha256:015a724affc8fa73361c4380433eca3be8fa1423f7327e98eb75ea48596444c7

Observation 2cef4d40-8a12-49c3-9d15-bc00610ca9e1 · outbound

This paper cites Pomo: Policy optimization with multiple optima for reinforcement learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Pomo: Policy optimization with multiple optima for reinforcement learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.457088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.457088Z digest=sha256:e564b53b83ea9020f3b66dcdaf6b490fb21a88920f78040f7e7c1c0e9ef4859f

Observation 80205747-990b-40b6-84b7-8ba9815b5e4e · outbound

This paper cites Mind dataset for diet planning and dietary healthcare with machine learning: dataset creation using combinatorial optimization and controllable generation with domain experts.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Mind dataset for diet planning and dietary healthcare with machine learning: dataset creation using combinatorial optimization and controllable generation with domain experts

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.176470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.461851Z digest=sha256:e781a3995a510fb4b88b51280ffc6f99154900ad8f25b3e0cf2f5efeea5269af

Observation a2ac4060-43e0-48b6-8bd0-2fb646b53ad6 · outbound

This paper cites Learning multi-level hierar- chies with hindsight.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Learning multi-level hierar- chies with hindsight

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.466597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.466597Z digest=sha256:91b5e2ac85ac307d5bc36ef18b8dfc347bcc47eb977902858c2380b49b7cdab9

Observation eec2e7cf-b65b-412b-a3d1-67f2ee418624 · outbound

This paper cites Deeper Insights Into Graph Convolutional Networks for Semi-Supervised Learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Deeper Insights Into Graph Convolutional Networks for Semi-Supervised Learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.149253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.471258Z digest=sha256:586a33fdfd8501a4760208634bcdf5b116d3473731084d9d8d5adcdb09af1680

Observation 53b0f20a-d133-472e-95af-d7c7763a832c · outbound

This paper cites Deep-learning-based wireless resource allocation with application to vehicular networks.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Deep-learning-based wireless resource allocation with application to vehicular networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.134149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.475492Z digest=sha256:2b1c8764e37e78c67bf534908f07bc7a46b4a94f7138af7eee1d7a2da28452f0

Observation bb161e73-031d-4ca3-97ce-e66c00828c30 · outbound

This paper cites Graph Foundation Models: Concepts, Opportunities and Challenges.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Graph Foundation Models: Concepts, Opportunities and Challenges

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.479770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.479770Z digest=sha256:c04ff9278cc20202abcfb5d6066a58605d4623d189c9ec8635f9bc6dbede2e9f

Observation 30db095e-8246-41c7-bd14-7d30d5f6072c · outbound

This paper cites A learning-based iterative method for solving vehicle routing problems.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning A learning-based iterative method for solving vehicle routing problems

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.119961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.484808Z digest=sha256:e3cce9ed0966f4f58ade95511437b9f8b2ac58cc59cb62876ce2b604dd31a07f

Observation 9a4472fc-995c-4090-96df-e06f811c33ba · outbound

This paper cites Reinforcement learning for combinatorial optimization: A survey.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Reinforcement learning for combinatorial optimization: A survey

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.105831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.488997Z digest=sha256:7391763e67efa0e47dafd0509d2efc7d1bb019e6a86d9c22bca6031145f63d9e

Observation 967d589c-0932-4ad7-811d-7b08b4d7210d · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.494362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.494362Z digest=sha256:ebc41740b4647046cb568b009e1501ab7258c6169a470fb64b71ccf62ecd7010

Observation 207edfba-cd35-45a7-9863-a94c18bb49a8 · outbound

This paper cites Human-level control through deep reinforcement learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Human-level control through deep reinforcement learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.499797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.499797Z digest=sha256:ea6064588d3d7bab3f3e8922a5950e9e5a7594bc90e5598a82827020c8efd586

Observation cf0b18ab-bd0d-429a-9660-4e91e7b6b0d7 · outbound

This paper cites Data-efficient hierarchical reinforcement learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Data-efficient hierarchical reinforcement learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.081050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.504359Z digest=sha256:71117b286c7d1b408059056daf994c20fce2e50c1805feb7e934a2f9da8c29cc

Observation 2ff9a79b-80bf-4460-8447-9033028d33bb · outbound

This paper cites Reinforcement learning for solving the vehicle routing problem.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Reinforcement learning for solving the vehicle routing problem

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.065490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.509865Z digest=sha256:79008a2949f27b3bbace2ec81620d8376f374d9fddd6fef5a20acade69f3c000

Observation 246a2440-7e7b-43bc-af6b-ae762afb04a0 · outbound

This paper cites Solo: search online, learn offline for combinatorial optimization problems.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Solo: search online, learn offline for combinatorial optimization problems

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.048989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.514510Z digest=sha256:c7edf8ff08e7d65cc462876ea39c304c7e59a82c7d2cbe4e574ea18181ecc04b

Observation 83b47402-5b46-4d48-ad9d-f6127206dc96 · outbound

This paper cites Active screening for recurrent diseases: A reinforcement learning approach.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Active screening for recurrent diseases: A reinforcement learning approach

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.033330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.526265Z digest=sha256:858e97cdbb161595ec7389129668525c7b4f205275d21a171bb5054d14f50f27

Observation 8b6c6c81-7f71-49d5-af77-7c80f94e8b0f · outbound

This paper cites Adaptive influence maximization with myopic feedback.Advances in Neural Information Processing Systems, 32, 2019.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Adaptive influence maximization with myopic feedback.Advances in Neural Information Processing Systems, 32, 2019

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.017390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.531358Z digest=sha256:17f642872e94356b62c5dd9dd60ca19e58fafa8da5121bdd1690e787847cc2e6

Observation 828aafdf-ef8f-4fef-89bf-07e24994a913 · outbound

This paper cites Deepwalk: Online learning of social repre- sentations.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Deepwalk: Online learning of social repre- sentations

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:06.002672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.536450Z digest=sha256:a4a787f3dc7001242abd7f2b1689e9ebb1b265999719ba8c12ee27455b3337b8

Observation 9a8c371d-4573-41fe-abc6-965b6d67feb0 · outbound

This paper cites Scarselli, M.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Scarselli, M

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.986934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.542776Z digest=sha256:9006e4f67709ceb49e3d12efc92a3793cafb117a84fccf83be7e468c77cef28b

Observation 3b84c475-f100-4b82-aca3-724f22386111 · outbound

This paper cites Combinatorial optimization with physics-inspired graph neural networks.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Combinatorial optimization with physics-inspired graph neural networks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.548651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.548651Z digest=sha256:e670805ae1b1d8df5bcf30e689415274b684463055f7826e8f5815399cca3d86

Observation d58d75d6-008c-4983-9f35-cf0c142f8d09 · outbound

This paper cites Learning to predict by the methods of temporal differences.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Learning to predict by the methods of temporal differences

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.958959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.553555Z digest=sha256:f6cad2943935dbc4794cd1d5d191a99a15cf9090e03a7a6a4947aa914a1d60aa

Observation 87d164f2-5c43-4c6c-854f-37a091e712aa · outbound

This paper cites Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.943619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.558800Z digest=sha256:2bc8e8db6375af2a34bc45acd24834e48e5ffeaff5ab7458e948f28f7771844b

Observation f83cec7f-51d8-46ce-bb53-7d15a57cd815 · outbound

This paper cites Time-constrained adaptive influence maximization.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Time-constrained adaptive influence maximization

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.926788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.562939Z digest=sha256:9715dc71cbe198d93e18685ccfbb8d2aa10d49f0254075cd39474380685a9b62

Observation 279574b4-c7a4-4f92-85b6-13b4b0a75da0 · outbound

This paper cites Graph attention networks.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Graph attention networks

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.908402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.567015Z digest=sha256:7435ef352ba64a230355e58a5d1bde02f895ce70d93677d737b076ebc7d2e10f

Observation d64edccd-f62f-4dcc-be02-b3ef8e8cd0e6 · outbound

This paper cites Feudal networks for hierarchical reinforcement learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Feudal networks for hierarchical reinforcement learning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.890658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.571246Z digest=sha256:5bc52e69d5f3fc73edf4623bdd6a670b583c628eb4b1e41179c66464d1a2e505

Observation 761b4a4b-41cc-4472-836b-2eae7c724760 · outbound

This paper cites Q-learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Q-learning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.873513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.576211Z digest=sha256:9c4f5554a3b1d65c878ee490b23503d7b7b901b94a027e553eb54f446aee704d

Observation 94f9e66f-b35b-41b9-8587-bfd7d718a5b7 · outbound

This paper cites On efficiency in hierarchical reinforcement learning.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning On efficiency in hierarchical reinforcement learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.857250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.581755Z digest=sha256:7f903a595e3687ead887df75c2074a0ae0724e8f850cc3c68d550ee72de7f2c0

Observation dcb0eca0-bb02-48f1-8da4-797527bdb221 · outbound

This paper cites an unresolved cited work.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-08T19:03:05.838649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.587774Z digest=sha256:df1f6dc73744fe96d3b03163c26ea466ba8c2e450c3db1c2db2e7b8a786b4b56

Observation 49ba5402-4cf3-4dd3-8839-9d705c867b04 · outbound

This paper cites How Powerful are Graph Neural Networks?.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning How Powerful are Graph Neural Networks?

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T19:03:05.592590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:03:05.592590Z digest=sha256:6ee7587e45ff3032712189304619d851901414281d1ca456a852068101cf2eb2

Observation 23ad6750-cbdc-4758-b9ac-cac85fe42033 · outbound

This paper cites Accelerating Exact Combinatorial Optimization via RL-based Ini- tialization - A Case Study in Scheduling.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Accelerating Exact Combinatorial Optimization via RL-based Ini- tialization - A Case Study in Scheduling

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.821556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.597802Z digest=sha256:9e8f3610e88e0bcb3db46cd9fcf4aefb3f5cc5d1694d358bb6fbe88b333390f0

Observation e364317d-dd16-435b-b6a0-fb87109b78ec · outbound

This paper cites Gnnexplainer: Generating explanations for graph neural networks.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Gnnexplainer: Generating explanations for graph neural networks

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.805142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.602875Z digest=sha256:1a28a7d8d575e93c263c3c2a8c57c20d54b9bece8e905495d7fb93412cbc40fe

Observation d38485cb-03f0-466e-b2e2-d4abaa383722 · outbound

This paper cites Graph transformer networks.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Graph transformer networks

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:03:05.788657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.607378Z digest=sha256:b940fb787b53a11ec285c1353004e31a5a5cea2214ed9d1ea5f0353cdd32bd33

Observation b8b130ab-4d76-4cf8-bdee-1637f0c75ed7 · outbound

This paper cites Graph neural networks: A review of methods and applications.

Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning Graph neural networks: A review of methods and applications

Reference 58

Resolution
malformed identifier
raw_fallback, observed 2026-08-08T19:03:05.773978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-08T19:03:05.612017Z digest=sha256:178ce74cf1d733e6d3b4da7e7499ba0ce2a84fef1c8f9153b232cc1d79273449

Pith citing papers

Observation 2a5b9ce3-0b75-478c-b9e1-ec527baf48a4 · inbound

Demonstrating Real Advantage of Machine-Learning-Enhanced Monte Carlo for Combinatorial Optimization cites this paper.

Demonstrating Real Advantage of Machine-Learning-Enhanced Monte Carlo for Combinatorial Optimization Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:55:54.392627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-18T04:53:10.425398Z digest=sha256:3cfb92f75d55a664adf2f2633840ea1d3ea7d43ebb6eba0bfa055a7c21532fb3