Pith. sign in

Paper Citation Record · LEDGER

Deep Reinforcement Learning in Parameterized Action Space

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:1511.04143.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1511.04143 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T11:20:19.980036Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T21:47:36.506263Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dc94fb86-bcb7-49ae-af24-eb4ae35ee822 · inbound

Learn A Flexible Exploration Model for Parameterized Action Markov Decision Processes cites this paper.

Learn A Flexible Exploration Model for Parameterized Action Markov Decision Processes Deep Reinforcement Learning in Parameterized Action Space

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T22:13:08.294832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:13:08.294832Z digest=sha256:7b8a9ee75fbf453a4df251ffd03a7ef36b297508c931b116f738d6e2c4de102d

Observation 9a3321a5-b8ff-4b77-8e96-fe4914d99042 · inbound

Expert-Free Online Transfer Learning in Multi-Agent Reinforcement Learning cites this paper.

Expert-Free Online Transfer Learning in Multi-Agent Reinforcement Learning Deep Reinforcement Learning in Parameterized Action Space

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T14:18:10.025780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:18:10.025780Z digest=sha256:2386d5bb7f997ea8b56c1abba4f89c0cc6b677c8b973c6a03eed4d58a4d7db25

Observation bdb8e2b7-50aa-4e68-88cb-a4ef82ea2765 · inbound

Towards Learning Scalable Agile Dynamic Motion Planning for Robosoccer Teams with Policy Optimization cites this paper.

Towards Learning Scalable Agile Dynamic Motion Planning for Robosoccer Teams with Policy Optimization Deep Reinforcement Learning in Parameterized Action Space

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T19:02:30.588787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:02:30.588787Z digest=sha256:9f7d3c8db3fb131dbfd9b76bc84dece3b6c2715412cbdbc76169c533170139db

Observation b1115345-e57e-4ec5-a187-8b9c1ed1f47e · inbound

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers cites this paper.

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers Deep Reinforcement Learning in Parameterized Action Space

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:07:23.686753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:07:23.686753Z digest=sha256:98b8ada0f1dfc7ce50c953733744cf05125e71fe6eb52000afe8fc2579c37490

Observation e336dbd1-6f49-48ba-9d56-3f05b482de7a · inbound

Hybrid TD3: Overestimation Bias Analysis and Stable Policy Optimization for Hybrid Action Space cites this paper.

Hybrid TD3: Overestimation Bias Analysis and Stable Policy Optimization for Hybrid Action Space Deep Reinforcement Learning in Parameterized Action Space

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T19:43:50.329677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:43:50.329677Z digest=sha256:ebb4c0d894c240b3c85d07d42b11c1a4f8ba296ca8db811853040b05e96269d4

Observation 8719416c-144c-4120-80ea-3edfd8c226a0 · inbound

Maturing Markov Decision Processes: Decision Making under Increasing Information and Shrinking Action Sets cites this paper.

Maturing Markov Decision Processes: Decision Making under Increasing Information and Shrinking Action Sets Deep Reinforcement Learning in Parameterized Action Space

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:49:02.360374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T21:47:16.490817Z digest=sha256:85da676b6f2b21248383d228fba2b1a385a7c46bc5eabd1425246b8dc67bd6f1

Observation 3cf8e4e4-6058-490a-bd89-2f555a4c812f · inbound

Integrated Automated Car Following and Lane-changing control based on a Parametrized Deep Q-network with Hybrid Action Space cites this paper.

Integrated Automated Car Following and Lane-changing control based on a Parametrized Deep Q-network with Hybrid Action Space Deep Reinforcement Learning in Parameterized Action Space

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T21:47:36.521773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-10T21:39:45.646371Z digest=sha256:ef316f75f53f2900995ff18d5c4b6a7d360a98ee15da4de0025009247d81a40b

Observation df969530-88a0-45bf-8ae8-eb2b1f682f6d · inbound

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition cites this paper.

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Deep Reinforcement Learning in Parameterized Action Space

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T11:20:19.980036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:20:19.980036Z digest=sha256:8b9967d8781efacdf68cfbbba155662a14b3adca90e95aeeb2e2c7b909e749a4