Pith. sign in

Paper Citation Record · LEDGER

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient

As of 10 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2507.09989.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09989 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:47:39.488108Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1fb07d7-507a-48af-9ee5-513978531797 · outbound

This paper cites Multi-Agent Reinforcement Learning for Power Control in Wireless Networks via Adaptive Graphs.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Multi-Agent Reinforcement Learning for Power Control in Wireless Networks via Adaptive Graphs

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:47:39.820541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:36.874336Z digest=sha256:0cb8bbf08d4c55b1a9cfac4f5329111ecf5b6cc7c7fd144c481a62b6cd9e36e5

Observation 8c506b9e-95db-4b15-b18b-0e8c687d69d7 · outbound

This paper cites Proceedings of the International Conference on Learning Representations (2022).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Proceedings of the International Conference on Learning Representations (2022)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.438220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:36.967492Z digest=sha256:7c7f95d19fda8403f9ba2188fe4d096f45e9a5c5354531d18e41482f6095f505

Observation 61012a7c-78bf-4851-98d2-78bf93e6df10 · outbound

This paper cites Pro- ceedings of the 2023 International Conference on Autonomous Agents and Multiagent Sys- tems pp.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Pro- ceedings of the 2023 International Conference on Autonomous Agents and Multiagent Sys- tems pp

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.422021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.010854Z digest=sha256:58e4fbd38483686e760aee0b9d3b2b1afe49bcbd515a7a7c53ad22783b49af81

Observation 78901d04-639e-469c-bbb3-88b198cb3e12 · outbound

This paper cites Joint European Conference on Machine Learning and Knowledge Discovery in Databases pp.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Joint European Conference on Machine Learning and Knowledge Discovery in Databases pp

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.403187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.093328Z digest=sha256:68f2278ebfc2e78ff0a1fda8323abe405e6e2d90670d84b24da343df7a8bde9b

Observation a28de0a2-a975-4e82-8050-52db1790c0dd · outbound

This paper cites International Joint Conference on Artificial Intelligence (2024).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient International Joint Conference on Artificial Intelligence (2024)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.377888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.169395Z digest=sha256:5f39b192ead72e78c8f52d5deb09d4dcfe1ab13f4417efd4680f1e8d76b8fffc

Observation b9f7b077-a641-4f1e-b536-eba274abbe39 · outbound

This paper cites Trust Region Policy Optimisation in Multi-Agent Reinforcement Learning.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Trust Region Policy Optimisation in Multi-Agent Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:37.242154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:37.242154Z digest=sha256:2abfc86300a61abaa32cebbf502d47eb4fce201ca2d7246cf17de5d59aa70d5e

Observation 063b6800-0521-46b1-85f8-6a051fef939d · outbound

This paper cites the Thirty-Eighth Annual Conference on Neural Information Pro- cessing Systems (NeurIPS) (2024).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient the Thirty-Eighth Annual Conference on Neural Information Pro- cessing Systems (NeurIPS) (2024)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.361427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.375135Z digest=sha256:485b0e1d89ff9705c1c408356c46ecac5671986fc36a7705814b9b479153a655

Observation 5d2937e1-04a3-4716-b108-7b0d664281a3 · outbound

This paper cites CoRR (2023).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient CoRR (2023)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.342816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.470113Z digest=sha256:591660aeae8176d8bc578803c50f298a5d64e809d21d57c6d7c205a694bc9e90

Observation 380c99e2-9524-4959-b545-1389d5e469e7 · outbound

This paper cites The Twelfth International Conference on Learning Representations (2024).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The Twelfth International Conference on Learning Representations (2024)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.323345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.537944Z digest=sha256:a4c7ad56ee61c1eb320ed9c8ee98d8e170f7f375c8a5b889f1c582a1e2566b21

Observation 7c7dd837-0402-40a1-aaeb-442128e50d50 · outbound

This paper cites The Twelfth International Conference on Learning Representations (2024) Title Suppressed Due to Excessive Length 13.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The Twelfth International Conference on Learning Representations (2024) Title Suppressed Due to Excessive Length 13

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.307050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.637960Z digest=sha256:d77a3ad4341bac4b86d81b3bc57d69cf7ff98cb4edba094b239570036d815210

Observation eb079445-1324-4aa5-9a0a-78e13167b24a · outbound

This paper cites Neural Information Processing Systems (NIPS) (2017).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Neural Information Processing Systems (NIPS) (2017)

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.289219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.696465Z digest=sha256:7eb54083371c5c68cf5bfa7baee6fa2c0c5a249ba4cc1d545090857ff5e5e76f

Observation fb8c8ab5-4842-438f-a72f-a58e8421c799 · outbound

This paper cites Advances in Neural Information Processing Systems 32 (2019).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Advances in Neural Information Processing Systems 32 (2019)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.270317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.781829Z digest=sha256:c83ebbd99c4894c4b78515a5b17977353f05da8a58e937702699c1b6766d14cd

Observation 439d077f-ba02-49f9-a46e-99ac3d29cc92 · outbound

This paper cites Springer (2016).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Springer (2016)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.249864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.882998Z digest=sha256:45c46ae795cf4d958f9338a46a0667cb825a5d60ecd312425f8c08f39f393189

Observation a7c27d62-01f7-4309-92ce-878c65b67101 · outbound

This paper cites Applied Intelligence 53(4), 4483–4498 (2023).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Applied Intelligence 53(4), 4483–4498 (2023)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.235075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:37.976283Z digest=sha256:43a6d012d7acf29cab74d81b7de1995840e13fba8298b82a7a3650c04687d966

Observation 76057f61-0aa5-49b1-b1de-78c6a887666c · outbound

This paper cites The Journal of Machine Learning Research 21(1), 7234–7284 (2020).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The Journal of Machine Learning Research 21(1), 7234–7284 (2020)

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.219283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:38.043046Z digest=sha256:c23891bd4be549da1f337b381ebbc6cb27c260ca9fec4c4337f4ede0330c16c8

Observation 6d3bb8c5-9eb2-482c-96b9-6d21f21b69eb · outbound

This paper cites FedMRL: Data Heterogeneity Aware Federated Multi-agent Deep Reinforcement Learning for Medical Imaging.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient FedMRL: Data Heterogeneity Aware Federated Multi-agent Deep Reinforcement Learning for Medical Imaging

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:38.095243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:38.095243Z digest=sha256:869b5cc7c6fc6a7d1ec40ddbde0d8c9a5e2a91cb7ec2402483d858544e1e36ed

Observation 3b5161f9-e1ee-498d-9936-389d9f3020d7 · outbound

This paper cites Neural Computing and Applications 35(27), 19765–19781 (2023).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Neural Computing and Applications 35(27), 19765–19781 (2023)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.199924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:38.208471Z digest=sha256:5b2dad5600e8eb5eb4aaf3a140311bc6fd41a435740fd43850771ccf49307ff1

Observation dffaa61c-db81-4f8f-81ab-fe11b132a75b · outbound

This paper cites Advances in Neural Information Processing Systems 35, 16509–16521 (2022).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Advances in Neural Information Processing Systems 35, 16509–16521 (2022)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.173531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:38.274554Z digest=sha256:888948ef85eb2e49bb75dcc53b77ca0fae1e751b08c02a19794f3bf41a8c288f

Observation 41ef0326-dcc2-416f-bc7c-779c1a4dc774 · outbound

This paper cites Theses and Dissertations.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Theses and Dissertations

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.148295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:38.399755Z digest=sha256:c6c343055464faba7635a0ba8b141ddf4e4c6d6cf5fb078baf89d046562508cc

Observation 63631aa8-cc7f-417e-939a-6a96e0d49dfa · outbound

This paper cites The International FLAIRS Conference Proceedings, 35 (2022).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The International FLAIRS Conference Proceedings, 35 (2022)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.124489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:38.542071Z digest=sha256:c3d44d0efeba964bc141dfd6ce46e23af7b4021b4bb1eef06e5609f6876f6e50

Observation 5e89f877-1b8c-44bc-8037-515e86d2fd20 · outbound

This paper cites IEEE Transactions on Vehicular Technology69(8), 8243–8256 (2020).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient IEEE Transactions on Vehicular Technology69(8), 8243–8256 (2020)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.102216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:38.596388Z digest=sha256:44c868b2aa2f90655a1f6fa73bf70d301200d4c4c4961ef8856f8d53e6aa1f04

Observation d5642dde-b820-4321-b5c4-1219b425b296 · outbound

This paper cites Designing Heterogeneous LLM Agents for Financial Sentiment Analysis.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Designing Heterogeneous LLM Agents for Financial Sentiment Analysis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:38.762403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:38.762403Z digest=sha256:97202e6a8ac871012ac74fa4d1970933195a9d0e263d2ca8b113c016bc7924f7

Observation 65dea1ea-b486-4dfd-8204-73e9e4c6c55a · outbound

This paper cites 2021 IEEE International Confer- ence on Systems, Man, and Cybernetics (SMC) pp.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient 2021 IEEE International Confer- ence on Systems, Man, and Cybernetics (SMC) pp

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.076193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:38.800514Z digest=sha256:6c1b2d89486cdf45e46ee64e57fcfe38830f4bd65d82f2e80511381d214d7993

Observation c86db028-91ff-41eb-89c9-caf595f0452f · outbound

This paper cites AIAA Scitech 2019 Forum p.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient AIAA Scitech 2019 Forum p

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.053893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:38.922343Z digest=sha256:0b67846ffd1277f900bd40603b6a9bf96a583e2875a514c018d6fb15219720b1

Observation 819dfc1c-9d73-4f90-98ea-28967e163a33 · outbound

This paper cites The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.086300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.086300Z digest=sha256:e685d23921881e3308bb9ef99ed32bea3ed5bbbf2eb705143bca2353189a6e4c

Observation a20098aa-3adb-4c03-aa7f-5b1476f29c16 · outbound

This paper cites Applied Sciences 15(5), 2580 (2025).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Applied Sciences 15(5), 2580 (2025)

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.032904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:39.148563Z digest=sha256:21341e748fe1267a086702981c84135a24132f995e5187890fc5fd8f028c25f5

Observation 42be7f25-7db7-4388-a20f-ffa157071d99 · outbound

This paper cites Complex & Intelligent Systems pp.

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Complex & Intelligent Systems pp

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:40.010555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:39.268295Z digest=sha256:3195d436e243d4762cc6a508cc7add5234fa760a28cedf13f6b78db3438f1139

Observation 6a71bfb5-a906-4ae5-9763-2a2ed60b78d7 · outbound

This paper cites Neurocomputing 411, 206–215 (2020).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Neurocomputing 411, 206–215 (2020)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:39.991068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:39.329041Z digest=sha256:9ca9d7285e37c3839b66b7a06bcaf276543f5a079b7d525e7760fe1dd738fb4f

Observation 998a9b7f-804e-4023-804a-b80a1f1d82d4 · outbound

This paper cites Autonomous Agents and Multi-Agent Systems 38(1), 4 (2024).

Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient Autonomous Agents and Multi-Agent Systems 38(1), 4 (2024)

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:47:39.969754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:47:39.488108Z digest=sha256:0e111731dbeecb9a5d7ab38b2d620d1d6d5d220974ccd9cbc5c67e83db8e9009

Pith citing papers

No inbound Pith citation observations are available.