Pith. sign in

Paper Citation Record · LEDGER

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking

As of 8 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2603.06607.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.06607 v2

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T22:34:20.119341Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9f1c05bb-9e6e-47b4-9ede-71541b8b4d0d · outbound

This paper cites Multi-agent drl for resource allocation in vehicular networks: A comparative study,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Multi-agent drl for resource allocation in vehicular networks: A comparative study,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.093903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.093903Z digest=sha256:685dda3ea7a21febf40fecc14d85074c0f089cd1dad91d783e0aae3aa4f941be

Observation 374ba54c-e49a-4153-9a02-3b7532b17dcb · outbound

This paper cites Recent advances and future trends in vehicular communication systems: A comprehensive survey,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Recent advances and future trends in vehicular communication systems: A comprehensive survey,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.216191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.216191Z digest=sha256:8d5a10c22913dee790f82096e970700cbfa88b32b81ed263597e8327759f52d5

Observation dd203d4d-334f-4b85-8187-2b9b1989a1cb · outbound

This paper cites Multiuser resource control with deep reinforcement learning in iot edge comput- ing,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Multiuser resource control with deep reinforcement learning in iot edge comput- ing,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.365192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.365192Z digest=sha256:ccf47f5e4c3a0d79487436b9e178cd77d45f1a65c9e68308f789e1f93be23ea1

Observation 4cc0e15b-acfc-4ae4-af43-c86f058257f7 · outbound

This paper cites Deep reinforcement learning based resource allocation for v2v communications,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Deep reinforcement learning based resource allocation for v2v communications,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.527546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.527546Z digest=sha256:d8cbe833ca2c7cc819de6b388667b55d70a8e18611752868756f27a3c1bb3c1d

Observation 106749c0-e2ce-4b2d-93cc-88533a1678de · outbound

This paper cites an unresolved cited work.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.599296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.599296Z digest=sha256:fd206244e79b934a9bf537449d65679c35613850e5ec2abd802863bc75900854

Observation 511f70b3-1e08-43aa-a109-56bc96060968 · outbound

This paper cites Spectrum sharing in vehicular networks based on multi-agent reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Spectrum sharing in vehicular networks based on multi-agent reinforcement learning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.661898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.661898Z digest=sha256:6a97a38b424e4eeffeb9097c0fe11799edde5dfdd162d6c6423bbbad9ec834ea

Observation fbb8b2fa-a04a-416b-a21e-4a6446d2c666 · outbound

This paper cites Multi-agent RL enables decentralized spectrum access in vehicular networks,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Multi-agent RL enables decentralized spectrum access in vehicular networks,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.707973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.707973Z digest=sha256:c0790c3148b6b6376a1179559731abcc927e8681aa0da764ee73fe4c1f7df6eb

Observation 3cb9e088-f01e-43b0-bf8b-e8502403d006 · outbound

This paper cites Multi-agent reinforcement learning-based decentralized spectrum access in vehic- ular networks with emergent communication,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Multi-agent reinforcement learning-based decentralized spectrum access in vehic- ular networks with emergent communication,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.766409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.766409Z digest=sha256:436c5725394c250d3b3c480bb061733c6f4fba2163c88ecb0d154f6646f96843

Observation 982bb0da-cf4c-480c-9155-70b5c1cc87b4 · outbound

This paper cites Resource allocation in V2X communications based on multi-agent reinforcement learning with attention mechanism,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Resource allocation in V2X communications based on multi-agent reinforcement learning with attention mechanism,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.825818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.825818Z digest=sha256:52ce876de9dc7e38cb12285e793b5657a617bf01fc420a55629a2cb56d36f4fb

Observation 74fef6dc-c209-4b86-a522-151d46cd5c62 · outbound

This paper cites Mean-field aided multi-agent reinforcement learning for resource allocation in vehicular networks,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Mean-field aided multi-agent reinforcement learning for resource allocation in vehicular networks,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.897576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.897576Z digest=sha256:2199d86b3711eb5c826dc051f1a35b8384aef010890c8acb5abc8086837b0614

Observation ac01e186-28d3-4e1f-ad11-5c21aa76f15e · outbound

This paper cites Graph Neural Networks and Deep Reinforcement Learning Based Resource Allocation for V2X Communications.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Graph Neural Networks and Deep Reinforcement Learning Based Resource Allocation for V2X Communications

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:16.986345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:16.986345Z digest=sha256:f908a1e541d4d1e71a7a7cbd7b4c36a5ac0bc8a0db5d0e2e87df590a3145e0d9

Observation 731b35d2-d8bf-493f-adc8-fe57b2b25cd8 · outbound

This paper cites Meta-reinforcement learning based resource allocation for dynamic v2x communications,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Meta-reinforcement learning based resource allocation for dynamic v2x communications,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.175140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.175140Z digest=sha256:1075464e94b8db5d37ddafb42eb1a68a6038121056b615920255f7c8d0f90e25

Observation 10b5bb66-4620-4565-8cc3-9b10e507eb35 · outbound

This paper cites Multitimescale control and communications with deep reinforcement learning—part ii: Control- aware radio resource allocation,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Multitimescale control and communications with deep reinforcement learning—part ii: Control- aware radio resource allocation,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.269895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.269895Z digest=sha256:f1ccd5916a2a4adcac658edeec0ee1c71af599cc9eaba2d1aa26f1242f588545

Observation 7963d358-bd87-40f6-b64f-5b9962f01b18 · outbound

This paper cites A review of cooperative multi-agent deep reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking A review of cooperative multi-agent deep reinforcement learning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.389777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.389777Z digest=sha256:4bc9f4c2dd995fefbecff37ac56bf8e640eb1857192bac34f48331f1abb027c4

Observation df126020-9847-4c4d-94d5-6dc1312ef6d7 · outbound

This paper cites Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.452139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.452139Z digest=sha256:8ef3335fc051de90862fc0f743dab82e94ca1f99e9aac37be58d36a353ce50d8

Observation 3cced00c-6889-4909-a1d6-e9c04a97e1bb · outbound

This paper cites Proximal Policy Optimization Algorithms.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Proximal Policy Optimization Algorithms

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.513122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.513122Z digest=sha256:ecf780675898135d6077ca540a23fdda4a9730601e141b09cce5d28d16387fdb

Observation 3320a770-529f-4dbc-b0a6-571027f34d67 · outbound

This paper cites Deep reinforcement learning based resource allocation with heterogeneous qos for cellular v2x,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Deep reinforcement learning based resource allocation with heterogeneous qos for cellular v2x,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.611716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.611716Z digest=sha256:ac068adf31e63752c9a9403bbda36a1222febaabf739ac5d79f93ae2469dc8f3

Observation 9a3779d8-374c-4d97-901a-1b71e5e04ff9 · outbound

This paper cites Spectrum-energy-efficient mode selection and resource allocation for heterogeneous v2x networks: A fed- erated multi-agent deep reinforcement learning approach,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Spectrum-energy-efficient mode selection and resource allocation for heterogeneous v2x networks: A fed- erated multi-agent deep reinforcement learning approach,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.673929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.673929Z digest=sha256:4a9408e05669ab1b262d5802a17e1948db0739039b29df0800d25efc33327d63

Observation b764cb70-958b-469f-bc43-f3fcb156e0ac · outbound

This paper cites A hybrid multi-agent reinforcement learning approach for spectrum sharing in vehicular networks,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking A hybrid multi-agent reinforcement learning approach for spectrum sharing in vehicular networks,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.746470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.746470Z digest=sha256:1e67a641df79333186260e33edeb752ab92a38e4b6c24d4ab83fce1a3da7d60f

Observation 4d9c1c0a-e034-42f3-95f8-ced5e59cf4bb · outbound

This paper cites Federated reinforcement learning for resource allocation in v2x networks,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Federated reinforcement learning for resource allocation in v2x networks,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.848633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.848633Z digest=sha256:6a09e2d7ef3a8096f5ae2d169822bba30dca3f837066d46697bcdf98bdcad068

Observation 49bb79f2-027a-42cc-a08c-18958143a13b · outbound

This paper cites Semantic- aware resource allocation based on deep reinforcement learning for 5g- v2x hetnets,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Semantic- aware resource allocation based on deep reinforcement learning for 5g- v2x hetnets,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:17.985066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:17.985066Z digest=sha256:13276672dd08d819020430e8d9642c4547538b7167b3b0e617afcb1a95d7b72f

Observation 7360ce8d-3795-4b10-8e6c-16d6fad6459b · outbound

This paper cites Deep reinforcement learning for multi- objective resource allocation in multi-platoon cooperative vehicular networks,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Deep reinforcement learning for multi- objective resource allocation in multi-platoon cooperative vehicular networks,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.064152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.064152Z digest=sha256:5a0d1d1ecfe850b1348b3dff3719d28da96c0567d7befaaa6bcc0b5f52124766

Observation 01bacd2b-7915-4c8d-a8e2-11ed06589450 · outbound

This paper cites Enabling adaptive optimization of energy efficiency and quality of service in nr-v2x communications via multiagent deep reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Enabling adaptive optimization of energy efficiency and quality of service in nr-v2x communications via multiagent deep reinforcement learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.179888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.179888Z digest=sha256:cf6ae16b8857ac98cf670df619be325c3b3cbfe061c35d422c8e100dcbcc9a54

Observation cf45f331-e31c-4e13-80a1-93390e7ff969 · outbound

This paper cites Aoi-aware resource allocation for platoon-based c-v2x networks via multi-agent multi-task reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Aoi-aware resource allocation for platoon-based c-v2x networks via multi-agent multi-task reinforcement learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.300005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.300005Z digest=sha256:79091c7b5a72c6a4e6e42590049b03075adc5d7babb9fbc97d8493c0e76e469c

Observation b0d3b2ca-6998-4be9-9d08-749536e122d2 · outbound

This paper cites Semantic-Aware Resource Management for C-V2X Platooning via Multi-Agent Reinforcement Learning.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Semantic-Aware Resource Management for C-V2X Platooning via Multi-Agent Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.388395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.388395Z digest=sha256:2833b0bd34a01eccb5342d96173c608fefb0e5bcc896cd6e45459dbaca0298d1

Observation 531f6618-ee62-4690-b966-43f7a7cc73d4 · outbound

This paper cites Deep reinforcement learning for autonomous internet of things: Model, ap- plications and challenges,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Deep reinforcement learning for autonomous internet of things: Model, ap- plications and challenges,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.491079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.491079Z digest=sha256:f93849c8791610e79c089e90e097cf10b3944e438ff37dcb1d40075187ec9113

Observation 74ad8ce7-7220-41e9-ad32-70887933cb06 · outbound

This paper cites Multi-agent actor-critic for mixed cooperative-competitive environ- ments,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Multi-agent actor-critic for mixed cooperative-competitive environ- ments,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.617606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.617606Z digest=sha256:caff10faee60c9f2c89a96a08685b12bdfccd548631190394770d798e43f6c86

Observation 0304c7ea-5099-4db4-8b3c-13e24a6d7f3a · outbound

This paper cites Continuous control with deep reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Continuous control with deep reinforcement learning,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.746179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.746179Z digest=sha256:dda4102380db499614ce13c79a92f2e60b398d3b4140797a2687e2793888eeb2

Observation 7f87d410-b139-483b-8a1c-ba2f563cd3b3 · outbound

This paper cites Technical Specification Group Radio Access Network; Study LTE- Based V2X Services; (Release 14),.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Technical Specification Group Radio Access Network; Study LTE- Based V2X Services; (Release 14),

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.826908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.826908Z digest=sha256:26f7aff08fa798bf3123374b578c6d4906d2e856e2acf5475cfe901a275520a9

Observation b9bb1224-a219-4aa6-a431-c24ecfb621e7 · outbound

This paper cites TR 103 766, no.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking TR 103 766, no

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:18.883970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:18.883970Z digest=sha256:c19816f83952b518afdabba813a045e51ea5e0fb9db1468bdb4ca4f60748b17f

Observation 17e50fc2-07fd-4da3-aedc-44c9e5989f8a · outbound

This paper cites Delay- optimal dynamic mode selection and resource allocation in device-to- device communications—part ii: Practical algorithm,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Delay- optimal dynamic mode selection and resource allocation in device-to- device communications—part ii: Practical algorithm,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.099121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.099121Z digest=sha256:14421577475b6d0bb0f53b1177f389da8de7c7b60442ca2c65d1fe89f4a93544

Observation ff5f13eb-31a7-42d6-8734-d8ed17dfd325 · outbound

This paper cites Lenient learning in independent-learner stochastic cooperative games,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Lenient learning in independent-learner stochastic cooperative games,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.227885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.227885Z digest=sha256:d40d364b238920f24092d8fb887bb427c955fe73fae6bed6c2a256efcbff4c24

Observation c4cd68c6-5569-40cd-b67b-0fbc82d50387 · outbound

This paper cites Leveraging procedural generation to benchmark reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Leveraging procedural generation to benchmark reinforcement learning,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.418454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.418454Z digest=sha256:76866bdff22073ddf8968136bf63eafdbe6ffa75df57429f8ffcb2dab44a9b10

Observation 3921ecd1-545a-474f-be8d-e190ede9d683 · outbound

This paper cites Human-level control through deep reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Human-level control through deep reinforcement learning,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.542941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.542941Z digest=sha256:3f6ccea6bc90c7ff7195b63ff4ca3d72b4c06317aa19b429b232dd22753dc061

Observation e406cb7e-4987-4d1d-9dc0-4edb0a6e72c0 · outbound

This paper cites Hysteretic q-learning: an algorithm for decentralized reinforcement learning in cooperative multi-agent teams,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Hysteretic q-learning: an algorithm for decentralized reinforcement learning in cooperative multi-agent teams,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.650964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.650964Z digest=sha256:4df68f9ca375025d4cbf779ee23cda7dabb374f39fb54b212c47eced56ab3335

Observation 2903a1ad-14ec-4221-acec-3c7a8e3cd478 · outbound

This paper cites Asynchronous methods for deep reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Asynchronous methods for deep reinforcement learning,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.700786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.700786Z digest=sha256:5ca8bafeb13ece28cde142719e5d62f1ec7a33e1098d7bd22b25e39f868cf5da

Observation d2be844e-5413-44bd-8d73-b66ede30eceb · outbound

This paper cites Value-decomposition networks for cooperative multi-agent learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Value-decomposition networks for cooperative multi-agent learning,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.775302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.775302Z digest=sha256:0f9b6297170269b96fa6a89dd62599119a1e3e3a1669d77e6a6de49e3347d28f

Observation e221153c-c6f8-404e-b5da-56b312ed82ad · outbound

This paper cites Qmix: Monotonic value function factorization for deep multi-agent reinforcement learning,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Qmix: Monotonic value function factorization for deep multi-agent reinforcement learning,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.853730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.853730Z digest=sha256:c8572fcf1ae25b4ffe045c734814a0030c59344d866642dce7cec2d5fa5dc955

Observation c355a984-4547-42ea-b39f-af769ff4e5ec · outbound

This paper cites The surprising effectiveness of ppo in cooperative, multi-agent games,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking The surprising effectiveness of ppo in cooperative, multi-agent games,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:19.938432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:19.938432Z digest=sha256:56d9ebd6f43b522d0108869434d0f525445b2e49f950f62cba3ddd4aa39586b9

Observation 37ad47fb-712a-4195-8c13-6ba7904afede · outbound

This paper cites Performance analysis of device-to-device communications with dynamic interference using stochastic petri nets,.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking Performance analysis of device-to-device communications with dynamic interference using stochastic petri nets,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:20.119341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:20.119341Z digest=sha256:dc7366e0aae7b4392c77189ae65e56c3e42cad9cbdf18b1b5662bc49a86de883

Observation c60c4bf8-78d6-4bb4-ae37-01190cc2feb3 · outbound

This paper cites The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games.

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T22:34:20.029978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:34:20.029978Z digest=sha256:2c735ec5a991d1d202d7a7a9cb614c1cf4afb106c64667e407789827c0687105

Pith citing papers

No inbound Pith citation observations are available.