Pith. sign in

Paper Citation Record · LEDGER

Mutual Information Tracks Policy Coherence in Reinforcement Learning

As of 17 August 2026, this Paper Citation Record lists 2 of 2 outbound references and 4 inbound Pith citation observations for arXiv:2509.10423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.10423 v1

Coverage vector

measured 2 of 2 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:58:32.377802Z

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T11:18:08.362562Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T11:24:38.222812Z

Reference resolution

2 of 2 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a245f6ea-f66a-4803-aa5e-853cdcb2d8a6 · outbound

This paper cites Information asymmetry in KL-regularized RL.

Mutual Information Tracks Policy Coherence in Reinforcement Learning Information asymmetry in KL-regularized RL

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:58:32.318597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:58:32.318597Z digest=sha256:0f303e38c8dd418eb9812af379ecf962ee7dc4ec768c16e50bae2fa07454eddd

Observation b5a832b8-54b9-4325-a76f-16d62b45ad20 · outbound

This paper cites The information bottleneck method.

Mutual Information Tracks Policy Coherence in Reinforcement Learning The information bottleneck method

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T15:58:32.377802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:58:32.377802Z digest=sha256:de2de482987a7a9e5a62cd18978d49ce4afb9c0fcfa54eb3fc042d8be6cdfa29

Pith citing papers

Observation 1966de2f-5f2e-4f59-abee-e7bcea1252dc · inbound

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning cites this paper.

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning Mutual Information Tracks Policy Coherence in Reinforcement Learning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:40:11.774606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-15T17:37:03.704563Z digest=sha256:f09e7668e634b54b805f77afc950d05265eaec63786dfd678d1567a3f4b27e25

Observation 476cde02-9fe0-40d1-95a5-a9795499325b · inbound

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning cites this paper.

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning Mutual Information Tracks Policy Coherence in Reinforcement Learning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:45:03.195195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-21T11:44:30.880602Z digest=sha256:8500679aa146e5409feb802ce0b261c4b4c708da8ec7dd584e1fcf744423636b

Observation b0201ee9-091a-46f7-b245-9426702b6bf7 · inbound

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty cites this paper.

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty Mutual Information Tracks Policy Coherence in Reinforcement Learning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T16:13:36.108563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T16:08:23.588699Z digest=sha256:c17477dcce54ca4e1445d1a12f3467c4b223adbc8d252d65e4a2ff5bf0f8e1fd

Observation 6cf2e075-683c-461b-aec7-7948e8030db3 · inbound

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty cites this paper.

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty Mutual Information Tracks Policy Coherence in Reinforcement Learning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:24:38.224205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T11:18:08.362562Z digest=sha256:d7652568564b73b27cc79d6da154533f47a6b46be23485d64c0fc2fdfb2a805c