Pith. sign in

Paper Citation Record · LEDGER

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient

As of 10 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2507.09989.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09989 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:47:39.488108Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1fb07d7-507a-48af-9ee5-513978531797 · outbound

This paper cites Multi-Agent Reinforcement Learning for Power Control in Wireless Networks via Adaptive Graphs.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Multi-Agent Reinforcement Learning for Power Control in Wireless Networks via Adaptive Graphs

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:47:39.820541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:36.874336Z digest=sha256:77820f329151d63c006dce9516850e7901b34433fbcfc3e16a96da5ff0043da0

Observation 8c506b9e-95db-4b15-b18b-0e8c687d69d7 · outbound

This paper cites Proceedings of the International Conference on Learning Representations (2022).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Proceedings of the International Conference on Learning Representations (2022)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.438220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:36.967492Z digest=sha256:8e902760bf549713df8029dd34df94c9773d56895f70202f137de24463964ca6

Observation 61012a7c-78bf-4851-98d2-78bf93e6df10 · outbound

This paper cites Pro- ceedings of the 2023 International Conference on Autonomous Agents and Multiagent Sys- tems pp.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Pro- ceedings of the 2023 International Conference on Autonomous Agents and Multiagent Sys- tems pp

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.422021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.010854Z digest=sha256:62acbd33ea412d9505e8140ff583b315659742ec78e0845f32fa452eb656506d

Observation 78901d04-639e-469c-bbb3-88b198cb3e12 · outbound

This paper cites Joint European Conference on Machine Learning and Knowledge Discovery in Databases pp.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Joint European Conference on Machine Learning and Knowledge Discovery in Databases pp

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.403187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.093328Z digest=sha256:9cd3a460d38aeca866f4691314d08101f5d13b09e9b28f1e6b8b1d478888e665

Observation a28de0a2-a975-4e82-8050-52db1790c0dd · outbound

This paper cites International Joint Conference on Artificial Intelligence (2024).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient International Joint Conference on Artificial Intelligence (2024)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.377888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.169395Z digest=sha256:2bcd117a6ea4497104f4a34deaca7b9c15d4ee115bbea1f1b5e4d3b75622a567

Observation b9f7b077-a641-4f1e-b536-eba274abbe39 · outbound

This paper cites Trust Region Policy Optimisation in Multi-Agent Reinforcement Learning.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Trust Region Policy Optimisation in Multi-Agent Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:37.242154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:37.242154Z digest=sha256:2abfc86300a61abaa32cebbf502d47eb4fce201ca2d7246cf17de5d59aa70d5e

Observation 063b6800-0521-46b1-85f8-6a051fef939d · outbound

This paper cites the Thirty-Eighth Annual Conference on Neural Information Pro- cessing Systems (NeurIPS) (2024).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient the Thirty-Eighth Annual Conference on Neural Information Pro- cessing Systems (NeurIPS) (2024)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.361427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.375135Z digest=sha256:63d9a6cdc95f7212430c5af3118003d0b1ae6fc7ab974fa4bcffaae1a5840d55

Observation 5d2937e1-04a3-4716-b108-7b0d664281a3 · outbound

This paper cites CoRR (2023).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient CoRR (2023)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.342816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.470113Z digest=sha256:a3422f3631f41bf045eef9b0554498ba8eb65c3d77a99987d2be22b5b79ff735

Observation 380c99e2-9524-4959-b545-1389d5e469e7 · outbound

This paper cites The Twelfth International Conference on Learning Representations (2024).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The Twelfth International Conference on Learning Representations (2024)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.323345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.537944Z digest=sha256:a4c0c4218f84ec4b6bc2af40a98096258d286d80c8f0c53efd14e2a3d9165474

Observation 7c7dd837-0402-40a1-aaeb-442128e50d50 · outbound

This paper cites The Twelfth International Conference on Learning Representations (2024) Title Suppressed Due to Excessive Length 13.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The Twelfth International Conference on Learning Representations (2024) Title Suppressed Due to Excessive Length 13

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.307050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.637960Z digest=sha256:ecf73c9c2ed23d4058e556085d17134a5abfb27c47b1a6831acaabd3f5bafef1

Observation eb079445-1324-4aa5-9a0a-78e13167b24a · outbound

This paper cites Neural Information Processing Systems (NIPS) (2017).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Neural Information Processing Systems (NIPS) (2017)

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.289219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.696465Z digest=sha256:326eb970688f68a8686b0d8b418afca10cd1f12d108e022212ecbce29fa9e9e4

Observation fb8c8ab5-4842-438f-a72f-a58e8421c799 · outbound

This paper cites Advances in Neural Information Processing Systems 32 (2019).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Advances in Neural Information Processing Systems 32 (2019)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.270317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.781829Z digest=sha256:fa80829d02a450e74c9d98cc69ca07d4ff0bc90f517863bd028e7cd767bf7eea

Observation 439d077f-ba02-49f9-a46e-99ac3d29cc92 · outbound

This paper cites Springer (2016).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Springer (2016)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.249864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.882998Z digest=sha256:3403971980ab186b216246d2d5953d3173d13d823582cd969811f69676f5f3f2

Observation a7c27d62-01f7-4309-92ce-878c65b67101 · outbound

This paper cites Applied Intelligence 53(4), 4483–4498 (2023).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Applied Intelligence 53(4), 4483–4498 (2023)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.235075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:37.976283Z digest=sha256:f310b1dba8163f4ab29db328887099239a22ee046599b714acd29dd95cf4e902

Observation 76057f61-0aa5-49b1-b1de-78c6a887666c · outbound

This paper cites The Journal of Machine Learning Research 21(1), 7234–7284 (2020).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The Journal of Machine Learning Research 21(1), 7234–7284 (2020)

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.219283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:38.043046Z digest=sha256:caed3a122e281a0cb6c3712d60f8e517afbb0bafb726bcdea2f781dfeb498772

Observation 6d3bb8c5-9eb2-482c-96b9-6d21f21b69eb · outbound

This paper cites FedMRL: Data Heterogeneity Aware Federated Multi-agent Deep Reinforcement Learning for Medical Imaging.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient FedMRL: Data Heterogeneity Aware Federated Multi-agent Deep Reinforcement Learning for Medical Imaging

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:38.095243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:38.095243Z digest=sha256:869b5cc7c6fc6a7d1ec40ddbde0d8c9a5e2a91cb7ec2402483d858544e1e36ed

Observation 3b5161f9-e1ee-498d-9936-389d9f3020d7 · outbound

This paper cites Neural Computing and Applications 35(27), 19765–19781 (2023).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Neural Computing and Applications 35(27), 19765–19781 (2023)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.199924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:38.208471Z digest=sha256:bb5bcbffd16c509ed85d8dec89480882524a70c420f6db47464944a513cb59a5

Observation dffaa61c-db81-4f8f-81ab-fe11b132a75b · outbound

This paper cites Advances in Neural Information Processing Systems 35, 16509–16521 (2022).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Advances in Neural Information Processing Systems 35, 16509–16521 (2022)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.173531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:38.274554Z digest=sha256:6e2cc8100f12a7aa92ea9b3ecf3fb2b29fc1d36208da7cddbab06df6fc13d463

Observation 41ef0326-dcc2-416f-bc7c-779c1a4dc774 · outbound

This paper cites Theses and Dissertations.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Theses and Dissertations

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.148295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:38.399755Z digest=sha256:9688d831afb38fae93f5ec50a9fde87644282d64fa46aef9b9b52b355c2bdad0

Observation 63631aa8-cc7f-417e-939a-6a96e0d49dfa · outbound

This paper cites The International FLAIRS Conference Proceedings, 35 (2022).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The International FLAIRS Conference Proceedings, 35 (2022)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.124489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:38.542071Z digest=sha256:69f7db08aadd4976ae2d9ce7f674bd9964b475b9d4012d9c8807cd3d2420c58f

Observation 5e89f877-1b8c-44bc-8037-515e86d2fd20 · outbound

This paper cites IEEE Transactions on Vehicular Technology69(8), 8243–8256 (2020).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient IEEE Transactions on Vehicular Technology69(8), 8243–8256 (2020)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.102216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:38.596388Z digest=sha256:1a2a85980e3c456915c941293519fca9569a1b3e2c4586b8b4072f8773d8a91c

Observation d5642dde-b820-4321-b5c4-1219b425b296 · outbound

This paper cites Designing Heterogeneous LLM Agents for Financial Sentiment Analysis.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Designing Heterogeneous LLM Agents for Financial Sentiment Analysis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:38.762403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:38.762403Z digest=sha256:97202e6a8ac871012ac74fa4d1970933195a9d0e263d2ca8b113c016bc7924f7

Observation 65dea1ea-b486-4dfd-8204-73e9e4c6c55a · outbound

This paper cites 2021 IEEE International Confer- ence on Systems, Man, and Cybernetics (SMC) pp.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient 2021 IEEE International Confer- ence on Systems, Man, and Cybernetics (SMC) pp

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.076193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:38.800514Z digest=sha256:582ef8d63332fc7f1c3e426bc7f90f46025e34e1aa5c8d244449943f735d917a

Observation c86db028-91ff-41eb-89c9-caf595f0452f · outbound

This paper cites AIAA Scitech 2019 Forum p.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient AIAA Scitech 2019 Forum p

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.053893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:38.922343Z digest=sha256:19f7f57e8099c0e952ac620dbbd3e4f2f4a9749c70c546ee1a7a92202791b04c

Observation 819dfc1c-9d73-4f90-98ea-28967e163a33 · outbound

This paper cites The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.086300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.086300Z digest=sha256:e08e2aef8e9f07bf37341a1cfce9fcc7dece42e6672ef82d1be765a8263aa47a

Observation a20098aa-3adb-4c03-aa7f-5b1476f29c16 · outbound

This paper cites Applied Sciences 15(5), 2580 (2025).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Applied Sciences 15(5), 2580 (2025)

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.032904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:39.148563Z digest=sha256:d16d100f260f28cafab87d6342de77517e1e69d5e1cf445efdf204484224f1f6

Observation 42be7f25-7db7-4388-a20f-ffa157071d99 · outbound

This paper cites Complex & Intelligent Systems pp.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Complex & Intelligent Systems pp

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.010555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:39.268295Z digest=sha256:b6814747784177972538d2bdc6ba33f463a6fea5432838de2c83320867336d43

Observation 6a71bfb5-a906-4ae5-9763-2a2ed60b78d7 · outbound

This paper cites Neurocomputing 411, 206–215 (2020).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Neurocomputing 411, 206–215 (2020)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:39.991068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:39.329041Z digest=sha256:364fade0a1ca8390a0c3e78b90fe820c3ca0c0440074e36831a050fd4f3548cc

Observation 998a9b7f-804e-4023-804a-b80a1f1d82d4 · outbound

This paper cites Autonomous Agents and Multi-Agent Systems 38(1), 4 (2024).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Autonomous Agents and Multi-Agent Systems 38(1), 4 (2024)

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:39.969754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T17:47:39.488108Z digest=sha256:fd22d5655db8b36c1cd4aa828009368f3192926184d50703b231549bd05ea091

Pith citing papers

No inbound Pith citation observations are available.