Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2006.14171.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:37:43.440877Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T08:39:41.546939Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 652b256a-eaaf-4233-94de-631fc1e5f332 · inbound
Effective Analog ICs Floorplanning with Relational Graph Neural Networks and Reinforcement Learning A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 810ba853-4e80-4603-a081-6ade4818f69b · inbound
Integrating Transit Signal Priority into Multi-Agent Reinforcement Learning based Traffic Signal Control A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c50f1d80-fa3f-41ac-af74-e2ad62556fc9 · inbound
Action Mapping for Reinforcement Learning in Continuous Environments with Constraints A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3793f95a-11f8-46cf-86c2-36f4bc144295 · inbound
Dynamic Collaborative Material Distribution System for Intelligent Robots In Smart Manufacturing A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53be11a9-015f-4c00-9348-8634fb2cef0c · inbound
Data-Driven Policy Mapping for Safe RL-based Energy Management Systems A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 116
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07e50d5f-b39a-42f3-8389-29f39ca7fc18 · inbound
Novel Multi-Agent Action Masked Deep Reinforcement Learning for General Industrial Assembly Lines Balancing Problems A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f748538-feec-4615-9928-e9c37290ccb1 · inbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbbb56f0-4485-47d6-b0bd-f13aa40aa6e1 · inbound
A Hierarchical Signal Coordination and Control System Using a Hybrid Model-based and Reinforcement Learning Approach A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0ad9eb3-73f8-444f-8359-c85db362407c · inbound
Learning to Assemble the Soma Cube with Legal-Action Masked DQN and Safe ZYZ Regrasp on a Doosan M0609 A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe0e8f53-bd8b-43a7-9692-8648a2b0ea01 · inbound
Towards Scalable O-RAN Resource Management: Graph-Augmented Proximal Policy Optimization A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b57996-e29b-451e-8ec2-88e49c4e411e · inbound
TARMM: Scaling Delay-Critical Edge AI Offloading in 5G O-RAN via Temporal Graph Mobility Management A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0a71177f-d0ee-4528-9be5-39d1d5221667 · inbound
Your Loss is My Gain: Low Stake Attacks on Liquid Staking Pools A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4f5e053d-38c2-49ea-9cfe-528aecb8588a · inbound
TuniQ: Autotuning Compilation Passes for Quantum Workloads at Scale for Effectiveness and Efficiency A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6a89b502-5468-4592-97a4-a75c77dbc6b8 · inbound
Learning Selective Merge Policies for Deadline-Constrained Coded Caching via Deep Reinforcement Learning A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7f968414-1195-4db5-977e-e03a6be706bb · inbound
Learning Selective Merge Policies for Deadline-Constrained Coded Caching via Deep Reinforcement Learning A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 67e32243-3306-42fd-971d-72c4e23371c6 · inbound
AlphaTransit: Learning to Design City-scale Transit Routes A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7293490f-2098-479d-9857-e34000b2ab40 · inbound
Bellman-Taylor Score Decoding for Markov Decision Processes with State-Dependent Feasible Action Sets A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a3538820-553e-4d4c-ae3c-5905c9db9c56 · inbound
Deep RL for Fast Long-Horizon Operations Scheduling on NASA's Carruthers Geocorona Observatory Mission A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 934b253b-b71e-4296-9722-fe10141e3034 · inbound
Optimal Reward Shaping: Autonomous Car Parking Case Study A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 244ead7a-3a31-461e-aa01-b6ae85d09cb4 · inbound
AlphaG-OPD: Reliability-Gated Sibling Counterfactuals for On-Policy Distillation in Symbolic Alpha Factor Discovery A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.