Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems

As of 18 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2501.06554.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.06554 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:02:18.838862Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact3
  • verified fuzzy15
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eadc3631-0158-434b-89ca-ae7084e7a4f3 · outbound

This paper cites The Option-Critic Architecture.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems The Option-Critic Architecture

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.674067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.674067Z digest=sha256:205fb6ce1640d0d5a6d1268e0f261ba51de8dde533fefb21f3e5450411793dcd

Observation 0fea0433-c780-4c95-9865-f9908bac73d3 · outbound

This paper cites Option-Critic in Cooperative Multi-agent Systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Option-Critic in Cooperative Multi-agent Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.679710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.679710Z digest=sha256:90fdf4a4a6823d3eff10acd8d6ce96acea68cafbdc292c91e69583517620b862

Observation f5ecb7a5-0afa-4c58-a1d6-afedcb1734fd · outbound

This paper cites Attention Option-Critic.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Attention Option-Critic

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.684622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.684622Z digest=sha256:f3754241203fa5c872a17c331cc426bb8b26f2d7ca48c946e8e88a85b4a0a807

Observation 95d6cceb-71fc-4029-b7cc-d00353e345d1 · outbound

This paper cites Dynamic planning in open-ended dialogue using reinforcement learning, 2022.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Dynamic planning in open-ended dialogue using reinforcement learning, 2022

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.303210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.694553Z digest=sha256:adf68070be4be404f83ae30565b6abd2b9631a1a1eb711701e60c99476a324d0

Observation 6c0a99da-a5ad-4ac9-a9e3-1e271965f52d · outbound

This paper cites Handbook on agent-oriented design processes.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Handbook on agent-oriented design processes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.277212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.701957Z digest=sha256:bf1163563047a610ddae6e11e84d1b952c1e4b069fbc93c34092159c392367f4

Observation 083f9d7f-3773-4241-bcaf-0371ef7eb66f · outbound

This paper cites Multi-agent deep reinforcement learning: a survey.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent deep reinforcement learning: a survey

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.258403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.707714Z digest=sha256:a7a048f394b3e2945a643c4dbc5e0887f11fa197fe6e451cec916777d52131e5

Observation c158bcc2-96ad-413d-a973-cabd98069dd6 · outbound

This paper cites Two-sided matching with firms' complementary preferences.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Two-sided matching with firms' complementary preferences

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:02:18.959877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.713199Z digest=sha256:0fc5d09cbf03ad72a06cceca4422dd8edf2330cc8b64c283dc236190309923d9

Observation b583ae01-4be9-4f64-a809-2756a559dd15 · outbound

This paper cites Emergence of division of labour in halictine bees: contributions of social interactions and behavioural variance.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Emergence of division of labour in halictine bees: contributions of social interactions and behavioural variance

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.235230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.722528Z digest=sha256:031a9815c3cb31572dc97272cf47fe5bd6a864031f4126c28787ed91a156f01c

Observation 83cdcd13-c859-4f58-88fb-bdaa6365972d · outbound

This paper cites Multi-agent deep reinforcement learning with type-based hierarchical group communication.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent deep reinforcement learning with type-based hierarchical group communication

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.215155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.729786Z digest=sha256:670db835f97a3c5a994af5381c3b32e568da9d1ef20556f96d364b92757ce382

Observation 090d9cbc-c6f9-4f6b-a59b-da3d7fd0a000 · outbound

This paper cites Multi-agent reinforcement learning as a rehearsal for decentralized planning.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Multi-agent reinforcement learning as a rehearsal for decentralized planning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.190680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.737992Z digest=sha256:9ab6ebe3ef0d672ee4044507867dc548b37ac4ca62f811592def20f49964f1e2

Observation 431133cc-3620-43dd-a3e7-2becf05f7e7b · outbound

This paper cites Reinforcement learning-based joint user pairing and power allocation in mimo-noma systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Reinforcement learning-based joint user pairing and power allocation in mimo-noma systems

Reference 11

Resolution
verified exact
doi, observed 2026-08-10T21:02:18.918637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.744243Z digest=sha256:9d92d0f3bb32fa34d702fb235ccebcd96cae6ac096d94c4f1bd9fb78c6bcda3e

Observation e289dd3b-78f5-4d0e-9483-30ed71682821 · outbound

This paper cites Role-based modeling for designing agent behavior in self-organizing multi-agent systems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Role-based modeling for designing agent behavior in self-organizing multi-agent systems

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.168675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.748639Z digest=sha256:0eccc4413937781290038cc4b95d008a81dcc06f788aca82a6103a0912142304

Observation b0aa34ba-ace1-4d61-bd3e-c2063dff96d8 · outbound

This paper cites Jordan, and Zhuoran Yang.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Jordan, and Zhuoran Yang

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.151523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.754362Z digest=sha256:a9ced42f99ae26a27fbb7d30b93397ff50e8141cbe074989709b97f97a72c32e

Observation c1121493-a712-4bd5-9e3b-a86a99abd94b · outbound

This paper cites Optimal and approximate q-value functions for decentralized pomdps.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Optimal and approximate q-value functions for decentralized pomdps

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.135232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.760712Z digest=sha256:20ddb89766c813815504096f2ed8828341806591ff8c81e9872b15e9b1f2db3b

Observation a4e77e8b-586f-4c33-a74f-83ea0aac2ac0 · outbound

This paper cites Hierarchical reinforcement learning: A comprehensive survey.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Hierarchical reinforcement learning: A comprehensive survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.768939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.768939Z digest=sha256:642dfbf2d55c3652facdf57a6e5f73fda8e4ddc6138fac85b209047a3584250b

Observation 58fb5ca5-523b-4bcc-8d02-a0972aa9a9da · outbound

This paper cites Vast: Value function factorization with variable agent sub-teams.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Vast: Value function factorization with variable agent sub-teams

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.776671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.776671Z digest=sha256:ab78a8df617500eed3d76a13c48e9d3ebc8a774e5182890fd5557ac4b76454ae

Observation ba9c940b-af4e-41bf-b55a-d52e52d2de95 · outbound

This paper cites Advances in neural information processing systems 17: proceedings of the 2004 conference, volume 17.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Advances in neural information processing systems 17: proceedings of the 2004 conference, volume 17

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.103251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.784316Z digest=sha256:7d7dd4646dde6c28ae3ba9b8584a96cb991e7960827c0042f182a9fb4f60c532

Observation 986987ca-1a48-4a50-810a-63112b531ac5 · outbound

This paper cites Sutton, Doina Precup, and Satinder Singh.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Sutton, Doina Precup, and Satinder Singh

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.796890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.796890Z digest=sha256:525d5b7126ed0b42022bc7f6fe38560f904a6830cc4c3c7f6e003e42954f868c

Observation c42a5f96-d91a-4728-b5e7-398ca9a73480 · outbound

This paper cites Effectiveness of gamified team competition as mhealth intervention for medical interns: a cluster micro-randomized trial.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Effectiveness of gamified team competition as mhealth intervention for medical interns: a cluster micro-randomized trial

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.084218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.802360Z digest=sha256:05c76dc1304cbf2c1d25f17648e5080cb4c554e8e9ca3f1613fff2f3d9f51cc2

Observation 25a0eddb-fe5e-48dd-b9dd-891f822cd454 · outbound

This paper cites Beaulieu, and Y.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Beaulieu, and Y

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.069932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.808135Z digest=sha256:58ea18c657245d1ccfbbd45d78205302c9b64975295d94a58574d52ea430b42c

Observation d7f37294-37b6-44e2-a88e-e1eb4c05c8e9 · outbound

This paper cites Adaptive dynamic bipartite graph matching: A reinforcement learning approach.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Adaptive dynamic bipartite graph matching: A reinforcement learning approach

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.055682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.813132Z digest=sha256:c39388924e61cd868e09480cb7369494021131005595e42d3aae325bee17a2be

Observation aa678bbc-d480-441f-a8c5-4eff6275f472 · outbound

This paper cites Hierarchical dominance structure and social organization in african elephants, loxodonta africana.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Hierarchical dominance structure and social organization in african elephants, loxodonta africana

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.036585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.821021Z digest=sha256:18f4b209067677ee667a78a19f147e7c776643729f1e6dcee7ffe18cc8b52776

Observation e7e48205-37c8-4852-b131-6c3c71240ba8 · outbound

This paper cites Large-scale order dispatch in on-demand ride-hailing platforms: A learning and planning approach.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Large-scale order dispatch in on-demand ride-hailing platforms: A learning and planning approach

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:19.022090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.826218Z digest=sha256:82820fa702d17d3b39307977ac47c2ba51b4c738b499ed1c1501d4a698c17507

Observation 8e214a0b-6254-44c9-a47d-0ce9e1dca286 · outbound

This paper cites Deep Sets.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Deep Sets

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:18.831276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:02:18.831276Z digest=sha256:4a2d4eb44d61d23398e9747fede4c161406411ab35991197f59903872629ff2c

Observation f0d3cc4c-a014-4406-b5e2-c81ee647b00b · outbound

This paper cites Reinforcement learning based local search for grouping problems.

Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems Reinforcement learning based local search for grouping problems

Reference 25

Resolution
verified exact
doi, observed 2026-08-10T21:02:18.895206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T21:02:18.838862Z digest=sha256:bb222c3bd1dbdf4a06c9e5201cf1074b958367dae677454e02ab28cef11d4aa4

Pith citing papers

No inbound Pith citation observations are available.