Pith. sign in

Paper Citation Record · LEDGER

Pragmatic Policy Development via Interpretable Behavior Cloning

As of 20 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2507.17056.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.17056 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:04:38.869043Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved7
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b0ab8493-e16b-49a2-ba25-e39cd2af6e56 · outbound

This paper cites Following (Matsson et al., 2024a), we restricted the analysis to classes of DMARDs rather than individ- ual drugs.

Pragmatic Policy Development via Interpretable Behavior Cloning Following (Matsson et al., 2024a), we restricted the analysis to classes of DMARDs rather than individ- ual drugs

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:04:38.961367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:04:38.863588Z digest=sha256:1781eea163e7e1d74bb5f30f18d53910eedb6a8c370826af341061cb4f859b09

Observation 1d254811-2815-4910-899a-99d12207bd4b · outbound

This paper cites The Doctor Just Won't Accept That!.

Pragmatic Policy Development via Interpretable Behavior Cloning The Doctor Just Won't Accept That!

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:04:38.842866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:04:38.842866Z digest=sha256:4e1ea6e43a591f2da593f17af8cbe089c3c4e8848f104bdf647a767b6bb5915a

Observation a7884742-40dd-4060-ab59-e0556f900d3f · outbound

This paper cites an unresolved cited work.

Pragmatic Policy Development via Interpretable Behavior Cloning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:04:38.953006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:04:38.866370Z digest=sha256:a336beee9d986e849a6a02a17a7b4c21c6e7708bfa6a41ba4b6b11332696e282

Observation 80477ffe-2319-429d-afda-65fcc8901c8c · outbound

This paper cites For the other methods—DQN, BCQ, and CQL— we used the implementation provided by Luo et al.

Pragmatic Policy Development via Interpretable Behavior Cloning For the other methods—DQN, BCQ, and CQL— we used the implementation provided by Luo et al

Reference 8

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T15:04:38.944333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:04:38.869043Z digest=sha256:9d8922c4012c5efd5fdc7bd7994764d68d49ad486af3059444623bf004029a03

Observation d409f753-1c02-4ca4-ab51-45312fdba3bd · outbound

This paper cites Behavioural cloning for driving robots over rough terrain.

Pragmatic Policy Development via Interpretable Behavior Cloning Behavioural cloning for driving robots over rough terrain

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:04:38.986896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:04:38.855301Z digest=sha256:f793455d85ebed1f9a26971e3dc848d1bcf7e708cbd63014f53c025492cf0b86

Observation d4132ec1-52da-45e4-ba46-67765db35e5c · outbound

This paper cites Conservative Q-Improvement: Reinforcement Learning for an Interpretable Decision-Tree Policy.

Pragmatic Policy Development via Interpretable Behavior Cloning Conservative Q-Improvement: Reinforcement Learning for an Interpretable Decision-Tree Policy

Reference 1983

Resolution
unresolved
no resolver link, observed 2026-08-06T15:04:38.852237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:04:38.852237Z digest=sha256:a263b688e6c73a9079fcf636304d7c9b871ba8cab673a7a5b9f2dc03d9682082

Observation 056f7f80-8196-40e6-867b-e795428858a7 · outbound

This paper cites Deep Reinforcement Learning for Sepsis Treatment.

Pragmatic Policy Development via Interpretable Behavior Cloning Deep Reinforcement Learning for Sepsis Treatment

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-06T15:04:38.849175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:04:38.849175Z digest=sha256:99122f272e02c699e45c1976e817b60b451cd2d8008c54541a24b93efcee9889

Observation 4641c43b-3c3a-47de-86f6-3ade251ed508 · outbound

This paper cites Distilling deep rein- forcement learning policies in soft decision trees.

Pragmatic Policy Development via Interpretable Behavior Cloning Distilling deep rein- forcement learning policies in soft decision trees

Reference 2014

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:04:39.003648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:04:38.829990Z digest=sha256:a3e782c6c9556f31faf2a67c3284b6db628b09436d8689989bd6ac54e0929196

Observation af3dd4ac-89ef-411e-840f-26d72d465e11 · outbound

This paper cites Smolen, Robert B.M.

Pragmatic Policy Development via Interpretable Behavior Cloning Smolen, Robert B.M

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:04:38.977980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:04:38.858179Z digest=sha256:eb9bbd3446f40c8e95c65b3aad423caabd4578ce5834a811971526adcd2415e1

Observation 27778344-32cf-4392-a7e1-e54db046e05b · outbound

This paper cites Off-policy deep reinforcement learning without ex- ploration.

Pragmatic Policy Development via Interpretable Behavior Cloning Off-policy deep reinforcement learning without ex- ploration

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:04:38.995325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:04:38.833477Z digest=sha256:660fdc3717dc530f8b52f67bdd899db76d7d5767e505fb97dc496cacfd057ef4

Observation 467fabfc-7175-4e64-8fb1-6d1efb5a2c63 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Pragmatic Policy Development via Interpretable Behavior Cloning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T15:04:38.839728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:04:38.839728Z digest=sha256:6cded33be465b84b96c2ac4bf54c635eb8a5f311a4f02c7c7295fee6f1a2a486

Observation b01d28c9-5d7b-4357-9db7-b5ef5a5cc74c · outbound

This paper cites Measuring Calibration in Deep Learning.

Pragmatic Policy Development via Interpretable Behavior Cloning Measuring Calibration in Deep Learning

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T15:04:38.845945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:04:38.845945Z digest=sha256:54ea0018f404cef8e0316578570705754573ec95e0f25ece4b61ffc667dc0ca6

Observation 8677122e-7a88-4289-ba20-b9c7e9dcdc09 · outbound

This paper cites Datasets Our experimental evaluation was based on two dis- tinct datasets related to the management of rheuma- toid arthritis (RA) and sepsis.

Pragmatic Policy Development via Interpretable Behavior Cloning Datasets Our experimental evaluation was based on two dis- tinct datasets related to the management of rheuma- toid arthritis (RA) and sepsis

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:04:38.970031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:04:38.861017Z digest=sha256:cb93c64894225fc52d8c0a29cfe3cfe0c6de084bf3594496efa64dfc6bfdc6a7

Observation 22e4a638-523a-46ca-bf1f-f1a50b3792db · outbound

This paper cites Evaluating Reinforcement Learning Algorithms in Observational Health Settings.

Pragmatic Policy Development via Interpretable Behavior Cloning Evaluating Reinforcement Learning Algorithms in Observational Health Settings

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T15:04:38.836400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:04:38.836400Z digest=sha256:fe543b8a2f1bf370bfac83773ab9ee7a292172cd41c126e8b8a75a03770d8bf9

Pith citing papers

No inbound Pith citation observations are available.