Pith. sign in

Paper Citation Record · LEDGER

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes

As of 19 August 2026, this Paper Citation Record lists 6 of 6 outbound references and 1 inbound Pith citation observation for arXiv:2512.14617.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.14617 v2

Coverage vector

measured 6 of 6 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:15:27.229715Z

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:43:00.121084Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-15T15:43:00.548193Z

Reference resolution

6 of 6 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 68411146-cfa5-4866-8cda-884838c60776 · outbound

This paper cites simulation-lemma.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes simulation-lemma

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:26.838755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:26.838755Z digest=sha256:5f5d5c9ec0038841277de30bdac23e6e5aaeacebe5b10fa7bbf9133042d83e07

Observation b9c49a03-2702-4384-8c3a-d6b0f4079b78 · outbound

This paper cites an unresolved cited work.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:27.021872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:27.021872Z digest=sha256:145bc1aa262b0c72243426e278603332e98ae1dcb53a1fbffb585d9909ead4b0

Observation fecd9376-00bc-4af6-8f2f-2ba42e1f7d61 · outbound

This paper cites error + V πt ¯M (b, q)−V ∗ ¯M (b, q) | {z } est.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes error + V πt ¯M (b, q)−V ∗ ¯M (b, q) | {z } est

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:27.229715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:27.229715Z digest=sha256:03c32c642739cf55bb431ae85cd7b7a85bb128e8e9f55b89a7f6d586d7cc5ced

Observation 706d583f-ad28-4a7d-ae7d-d56150517726 · outbound

This paper cites Near-optimal Reinforcement Learning in Factored MDPs.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes Near-optimal Reinforcement Learning in Factored MDPs

Reference 2003

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:26.717241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:26.717241Z digest=sha256:3c4fd6eef090abc1a098b853be25929ff23ca6f83f3bb4f5ec137d0c58f95130

Observation 4bfe4a1d-56f7-4b16-a973-fa577001f1b7 · outbound

This paper cites InProc, ICAPS, volume 34, 13659–13662.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes InProc, ICAPS, volume 34, 13659–13662

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:26.591251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:26.591251Z digest=sha256:229aea216849693f54943f93d32c2a2c4273e5db64215d509a574ebe6228afa5

Observation a0d59e86-7840-4126-bf1b-495cf87f8e6e · outbound

This paper cites InInternational Conference on Artificial Intelligence and Statistics (AISTATS), 4114–4146.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes InInternational Conference on Artificial Intelligence and Statistics (AISTATS), 4114–4146

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:26.476410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:26.476410Z digest=sha256:f287053fe2972b77063a4b4dea9dfa51f27b1b82a02bd9891b493567a332c055

Pith citing papers

Observation a3287bdd-5ede-4ab7-bf89-b864090bc05b · inbound

Theoretical Foundations of $\max$@$k$ Reinforcement Learning cites this paper.

Theoretical Foundations of $\max$@$k$ Reinforcement Learning Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-08-15T15:43:00.554958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:43:00.121084Z digest=sha256:3e97d401dc9f4072fd4a5eb70349725ba5223744eb79a3bcb71156826582a37b