Pith. sign in

Paper Citation Record · LEDGER

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain

As of 17 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2506.06786.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06786 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:55:08.160720Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7398286a-4956-4f95-a711-1972aa3f39ce · outbound

This paper cites Exploration in deep reinforcement learning: A survey,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Exploration in deep reinforcement learning: A survey,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.037952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.037952Z digest=sha256:a652c9207a22ad704e214780db26c545f3289d18a9ad92b7699b634dd8768342

Observation 2b7abad1-1726-4681-be06-c2e94e5c61fb · outbound

This paper cites Deep reinforcement learning for time-critical wilderness search and rescue using drones,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Deep reinforcement learning for time-critical wilderness search and rescue using drones,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.043659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.043659Z digest=sha256:fa1e3bc35434d3e72d101b8c4237144ea36f9c6eca0a401db4373fb4d81a4b6a

Observation 292fd82b-4056-49fd-bcbe-7899350d43ac · outbound

This paper cites MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.048563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.048563Z digest=sha256:4d32f2646bf08745e604e6aca9037fa627bd02ef1630496cdc0fb5ed1478d4f8

Observation e5e0854f-babb-4008-85b6-8929849cc418 · outbound

This paper cites Regret bounds for information-directed reinforcement learning,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Regret bounds for information-directed reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.540988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.054122Z digest=sha256:6a1e8d78574ec225e5c5f20b2ca050688a67a8b0cc4ee632357eb49f3130d812

Observation adf17a1d-4b14-4273-a5f1-24b4f5ddc7b7 · outbound

This paper cites Selective exploration and information gathering in search and rescue using hierarchical learning guided by natural language input,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Selective exploration and information gathering in search and rescue using hierarchical learning guided by natural language input,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.523642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.059018Z digest=sha256:07f933a458c0e0c1ea89e1f89301be58fc39975528be8b15494218307d3a285b

Observation 021747fd-36ab-4553-82ce-91559fef86b9 · outbound

This paper cites Adversar: Adversarial search and rescue via multi-agent reinforcement learning,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Adversar: Adversarial search and rescue via multi-agent reinforcement learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.506480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.064858Z digest=sha256:78c2bb29cfd14709f12fc7e23fae5c061d02c2924fc31e49379eba50a5bd4a14

Observation 7867432d-7291-4cf8-8fbc-71340a251fdd · outbound

This paper cites Target search and navigation in heterogeneous robot systems with deep reinforcement learning,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Target search and navigation in heterogeneous robot systems with deep reinforcement learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.490184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.070670Z digest=sha256:61451525fd4745f49e35371ff9202341562e2ffd0078764d4c50c3b346516c69

Observation f7627ad9-0862-4d8c-8e21-2015ab9d841f · outbound

This paper cites Multi-robot cooperative target search based on distributed reinforcement learning method in 3d dynamic environments,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Multi-robot cooperative target search based on distributed reinforcement learning method in 3d dynamic environments,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.473171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.075594Z digest=sha256:207f1b1769877ba51321f748181cd5aeab471febb2df16e686949a812f2dcafc

Observation 399349a3-5f65-4f6e-8ac5-275f1bafbadf · outbound

This paper cites an unresolved cited work.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.080420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.080420Z digest=sha256:5b79564451476e667b21366d2af97a8e1a7c4ce4130612693ce222cb5d920210

Observation 9d35c641-da1d-4d39-933d-4eb8530a0265 · outbound

This paper cites Boltzmann exploration done right,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Boltzmann exploration done right,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.444604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.085052Z digest=sha256:4d282c0d7bf4fcb5a7022dc9b2934e231d715207b874eb2dde0c037a7a9963ca

Observation 35cbcc6b-197f-486b-9597-9c78e7c6b97d · outbound

This paper cites Using confidence bounds for exploitation-exploration trade- offs,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Using confidence bounds for exploitation-exploration trade- offs,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.426789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.089768Z digest=sha256:379a019e5952a9a18e81d4820e8cd28523d45582877cd1971ef2b28d5bd3046f

Observation 4a812954-355a-49bf-a755-6be8287c776d · outbound

This paper cites An empirical evaluation of thompson sam- pling,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain An empirical evaluation of thompson sam- pling,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.408573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.094471Z digest=sha256:3f29da0696d1211a7c31b59294a27d6d592c2b59a779e331f7f7a26298f5ccbc

Observation b9813539-f0d5-48f3-8c5c-cada4c50e4c9 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Curiosity-driven exploration by self-supervised prediction,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.099735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.099735Z digest=sha256:5813bb3b2753e101bc85e8d8a41fa8affef9fb1f86a292fe4fcd26a87a40d02c

Observation 95e244b5-9f07-4d80-b95f-e0a892e61399 · outbound

This paper cites Unifying count-based exploration and intrinsic motiva- tion,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Unifying count-based exploration and intrinsic motiva- tion,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.104666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.104666Z digest=sha256:6d2dd21cbc012495bc9efc2ae35f6cd2ceadd2c5d525e6fdcda07cf357947e1f

Observation 1e91654f-7996-496d-9846-07d9ccb8f408 · outbound

This paper cites Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.110056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.110056Z digest=sha256:6999351ea013a37381c90495dc2b8d0a6c0426d1166ca140cf5d861d8038ba8e

Observation 217eb0e2-bdb2-4011-aec0-5feafce2fd0a · outbound

This paper cites Exploration by Random Network Distillation.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Exploration by Random Network Distillation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.116114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.116114Z digest=sha256:85e146c01d8622ec4b8a9edc940fffdec7060dbd50ea89205ba96b75ef6166bb

Observation 20bf38b9-0336-400e-80d1-e7f705d683df · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Model-agnostic meta-learning for fast adaptation of deep networks,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.121560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.121560Z digest=sha256:2cf839ec9bfa2c9680119c506eee79c99c997d7fb4c46b0c7245475a5c36f3fc

Observation beca5d45-bc1d-40dd-93c9-efe68b3f5075 · outbound

This paper cites Lattimore and C.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Lattimore and C

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.358828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.126666Z digest=sha256:4a8d7af8a00f558ec2ec808a34ba4218a80d8df33027718d9470214e70113a49

Observation d2b74e39-e08d-469b-bce2-248c23810283 · outbound

This paper cites Hidden parameter markov decision processes: A semiparametric regression approach for discovering latent task parametrizations,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Hidden parameter markov decision processes: A semiparametric regression approach for discovering latent task parametrizations,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.341394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.131798Z digest=sha256:50f37027d312d7e0368e7b49f6f94d8d68a24e63506cdb5d021b303263733168

Observation 9135ebbd-6a7a-4ac0-89ff-ded3ac6cd182 · outbound

This paper cites Contextual Markov Decision Processes.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Contextual Markov Decision Processes

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.137170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.137170Z digest=sha256:e39d8a558e3682453447bd020765d9160b9cbc8bfa8b1d0c4a69264ed2f98d7d

Observation 5d9ee41b-2377-4147-b8d2-58f513e91a1e · outbound

This paper cites Reinforcement Learning in Presence of Discrete Markovian Context Evolution.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Reinforcement Learning in Presence of Discrete Markovian Context Evolution

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:55:08.207365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.142795Z digest=sha256:64ff1920be10793df6b84f9de8bbae36fecdab2158fb7c76a5affeb58191e2b6

Observation 3e2e32e3-2288-4089-91a9-9d2c56dae461 · outbound

This paper cites Context-aware dynamics model for generalization in model-based reinforcement learning,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Context-aware dynamics model for generalization in model-based reinforcement learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.323520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.149028Z digest=sha256:1d9d68277abceb905ca748f9377f3412b5449d618aca7f9fac92035acf59915d

Observation 26648876-6b35-45f1-8482-0234814bf0cc · outbound

This paper cites Learning to optimize via information- directed sampling,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Learning to optimize via information- directed sampling,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:55:08.306067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:55:08.155554Z digest=sha256:7e71b4b9c28be80490eb9b0e1a8e01714d7a7c27484d34d8f8084a66efb13722

Observation 7f08c08f-3949-43d9-b9a4-7393359e6b6f · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:55:08.160720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:55:08.160720Z digest=sha256:d38f5cdb948ff1ead763333286535149bb0ff62924fa33df094acb75d4a12adb

Pith citing papers

No inbound Pith citation observations are available.