Pith. sign in

Paper Citation Record · LEDGER

Model-Based Reinforcement Learning via Meta-Policy Optimization

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:1809.05214.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1809.05214 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:42:10.205189Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-25T20:16:11.737123Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a680c906-36f4-490b-b9a5-71c21e1a9d83 · inbound

Calibrated Model-Based Deep Reinforcement Learning cites this paper.

Calibrated Model-Based Deep Reinforcement Learning Model-Based Reinforcement Learning via Meta-Policy Optimization

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-25T20:16:11.742093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-25T20:13:35.127716Z digest=sha256:fc26d4baa506571f6a949d503b0c2dfbf36be63c3927fee621a1f115240ccd87

Observation 22df83cc-42e1-4b71-bea4-1a65449304d5 · inbound

Uncertainty-aware Model-based Policy Optimization cites this paper.

Uncertainty-aware Model-based Policy Optimization Model-Based Reinforcement Learning via Meta-Policy Optimization

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-25T16:17:03.951909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T16:16:29.296376Z digest=sha256:e95732c7bb40f347983ab254fbceab4ec0f3a191cf1da89eceee8097673cd3ea

Observation 3f3c85e4-827e-4691-bf16-4d1eb625a66b · inbound

Benchmarking Model-Based Reinforcement Learning cites this paper.

Benchmarking Model-Based Reinforcement Learning Model-Based Reinforcement Learning via Meta-Policy Optimization

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-25T10:15:37.139369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T10:12:49.348162Z digest=sha256:bc9bdfcd284bf2b34830f278648f542f5f56a81100ac4306f10986c6fd8306d4

Observation b1cb7d9f-9616-4e94-b017-47126842144f · inbound

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation cites this paper.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Model-Based Reinforcement Learning via Meta-Policy Optimization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.205189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.205189Z digest=sha256:88fa62751925250c53be650d7a56492d975228822edbb6d2e1d6578bd41a436b