Pith. sign in

Paper Citation Record · LEDGER

Q-Learning for Continuous Actions with Cross-Entropy Guided Policies

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 2 inbound Pith citation observations for arXiv:1903.10605.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1903.10605 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 2 of 2 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:28:38.396922Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T04:39:11.650771Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d14fd665-77f4-409e-8587-75aaf4caee95 · inbound

Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone cites this paper.

Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone Q-Learning for Continuous Actions with Cross-Entropy Guided Policies

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T19:28:38.396922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:28:38.396922Z digest=sha256:0b16dd92a81b4a6025e45bfe60acc89c79a39518c08618234430bfd2c37d1326

Observation 04bfc398-424b-4cec-927b-236058396c63 · inbound

Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies cites this paper.

Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies Q-Learning for Continuous Actions with Cross-Entropy Guided Policies

Reference 46

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T04:39:11.817965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T04:39:07.677891Z digest=sha256:153d2e8a4a8c557b909452c2782c8cf42c908bbae3e03eac4a4a1c7a42dfb6d7