Pith. sign in

Paper Citation Record · LEDGER

Deep Reinforcement Learning in Parameterized Action Space

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:1511.04143.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1511.04143 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:07:23.686753Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T21:47:36.506263Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b1115345-e57e-4ec5-a187-8b9c1ed1f47e · inbound

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers cites this paper.

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers Deep Reinforcement Learning in Parameterized Action Space

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:07:23.686753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:07:23.686753Z digest=sha256:265695e3d06c8dad3b230132ea6c79a29eccf796a323a18e46ccaed785c4cc24

Observation e336dbd1-6f49-48ba-9d56-3f05b482de7a · inbound

Hybrid TD3: Overestimation Bias Analysis and Stable Policy Optimization for Hybrid Action Space cites this paper.

Hybrid TD3: Overestimation Bias Analysis and Stable Policy Optimization for Hybrid Action Space Deep Reinforcement Learning in Parameterized Action Space

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T19:43:50.329677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:43:50.329677Z digest=sha256:23478708c5aef92fd21b53b5802d00010c0388209fa90cfe45e7f92ab3f00e07

Observation 8719416c-144c-4120-80ea-3edfd8c226a0 · inbound

Maturing Markov Decision Processes: Decision Making under Increasing Information and Shrinking Action Sets cites this paper.

Maturing Markov Decision Processes: Decision Making under Increasing Information and Shrinking Action Sets Deep Reinforcement Learning in Parameterized Action Space

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:49:02.360374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T21:47:16.490817Z digest=sha256:52d098d042aa4e547fa455637b5425aa062fa4f0c5ea3e013c600c7bd0936994

Observation 3cf8e4e4-6058-490a-bd89-2f555a4c812f · inbound

Integrated Automated Car Following and Lane-changing control based on a Parametrized Deep Q-network with Hybrid Action Space cites this paper.

Integrated Automated Car Following and Lane-changing control based on a Parametrized Deep Q-network with Hybrid Action Space Deep Reinforcement Learning in Parameterized Action Space

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T21:47:36.521773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-10T21:39:45.646371Z digest=sha256:1a0692783e485d2393347d5f649133f273bbcdb17bc48036cc73a307a6f532d9