Pith. sign in

REVIEW 2 cited by

Offline and Distributional Reinforcement Learning for Radio Resource Management

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.16764 v2 pith:AS447S6F submitted 2024-09-25 cs.LG cs.AIcs.MA

classification cs.LGcs.AIcs.MA
keywords onlineinteractionmanagementofflineresourceschemeadditiondistributional
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Reinforcement learning (RL) has proved to have a promising role in future intelligent wireless networks. Online RL has been adopted for radio resource management (RRM), taking over traditional schemes. However, due to its reliance on online interaction with the environment, its role becomes limited in practical, real-world problems where online interaction is not feasible. In addition, traditional RL stands short in front of the uncertainties and risks in real-world stochastic environments. In this manner, we propose an offline and distributional RL scheme for the RRM problem, enabling offline training using a static dataset without any interaction with the environment and considering the sources of uncertainties using the distributions of the return. Simulation results demonstrate that the proposed scheme outperforms conventional resource management models. In addition, it is the only scheme that surpasses online RL with a 10 % gain over online RL.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Resilient UAV Trajectory Planning via Few-Shot Meta-Offline Reinforcement Learning

    cs.RO 2025-02 conditional novelty 4.0 of 10

    A hybrid meta-offline reinforcement learning algorithm trains a UAV to minimize data age and transmission power from static datasets and adapts to new tasks in under 40 epochs.

  2. An Offline Multi-Agent Reinforcement Learning Framework for Radio Resource Management

    cs.MA 2025-01 conditional novelty 4.0 of 10

    Conservative Q-learning with offline multi-agent training improves simulated radio resource scheduling, and centralized training with decentralized execution gives the best complexity-performance trade-off.

Pith tools