Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning with Targeted Causal Interventions

As of 18 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2507.04373.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04373 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:59:01.602504Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved8
  • parse uncertain2
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bcd355d3-c603-4b23-a5b9-f8a40174012e · outbound

This paper cites For” loop, it records a trajectory τj. The “While.

Hierarchical Reinforcement Learning with Targeted Causal Interventions For” loop, it records a trajectory τj. The “While

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:06.021733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:58:59.566592Z digest=sha256:1a11f9419c7435536ba75480cf0c8d3a8d30308f34d37dcc335a1efab6307cb1

Observation 95ea9249-855c-411c-a23e-662050c6afee · outbound

This paper cites As we mentioned earlier, q1 is the upper bound on the expected number of the nodes that are ancestors of the node n.

Hierarchical Reinforcement Learning with Targeted Causal Interventions As we mentioned earlier, q1 is the upper bound on the expected number of the nodes that are ancestors of the node n

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.944378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:58:59.618392Z digest=sha256:aefc68096adc8a54e1d80c2c486beaa9cfe81a7aa18c2e84a58c0c04c61c8b92

Observation 69954592-9938-4d56-bf2e-2c8dc644aadb · outbound

This paper cites At this point, we have constructed Ijr.

Hierarchical Reinforcement Learning with Targeted Causal Interventions At this point, we have constructed Ijr

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.518464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:58:59.825900Z digest=sha256:8d0458581f7084d831a5ffd0d9b7143f73de5e25aa2e217f744cb9c10a375522

Observation 3c607cc7-9191-47d2-aed3-1feccd0c5f42 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.057351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:00.670188Z digest=sha256:f4a781ab2533e576b5b3da9a9cf4735a260aae598dac57d4b80bb8e0fb36491e

Observation edde5922-6480-4b98-ab29-786f37d13290 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:05.828947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:58:59.661189Z digest=sha256:62ab4eab47ce4136b823108ffd232fc92d91cc6968b0c8c11bf8f567cd29cd76

Observation 42e29f81-69a3-43a8-b12a-03053e05786f · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:05.693361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:58:59.724253Z digest=sha256:5d47dec8b69d8a09aba4f240ee0097ba54e610f93f2e0bbdb172a49f51c9acd7

Observation fa4077de-a00f-4075-9185-2e15929efd7f · outbound

This paper cites We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi.

Hierarchical Reinforcement Learning with Targeted Causal Interventions We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.217088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:58:59.911931Z digest=sha256:3d5164ae78c14b227b88291a128665f34eeb404663e52927db4f3c6d01c2cd0d

Observation 9e8f781a-6a32-4a70-8290-de481f5cb137 · outbound

This paper cites • If gi is the goal, backtrack to construct the path.

Hierarchical Reinforcement Learning with Targeted Causal Interventions • If gi is the goal, backtrack to construct the path

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.993906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:00.025603Z digest=sha256:ce97f2f7d854c7eb252f80109addf253de22fa7d79bc1e39795d48f8f2e7a37a

Observation 0fdd8088-47d8-400b-8192-c94b8819f64e · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.830033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:00.125185Z digest=sha256:c55dae1bbf9186bbeca366b050dbb99b76f410f1e7e92bd09d84b642f6e7ebd8

Observation 1d7564ee-af70-4be3-a534-1afc40a97600 · outbound

This paper cites We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi.

Hierarchical Reinforcement Learning with Targeted Causal Interventions We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.636736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:00.267033Z digest=sha256:c97a79a0b18df4b649ff98757e5bae477b2afc790d2ebd35fe652328cf69ae5d

Observation 85b7b82b-c5fa-4cc4-a2ce-39d18e44cbd0 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.475273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:00.396920Z digest=sha256:3f64f012fc3542e868cd28f9d4546cfaaae845927d039fae571b9da71bab263b

Observation 6a16f74e-6055-4dfa-af95-128e97919592 · outbound

This paper cites – For each neighbor gv of gu in ˆGt: * Set w(gu, gv) = |C HˆGt gu \ back track(gu)| + 1.

Hierarchical Reinforcement Learning with Targeted Causal Interventions – For each neighbor gv of gu in ˆGt: * Set w(gu, gv) = |C HˆGt gu \ back track(gu)| + 1

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.279653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:00.547872Z digest=sha256:31ae533014a19ccb76daa0e37f1f2b3066c2e4fe2b9f0416a243c99ffda899d3

Observation 8eb58f2e-b6e1-4739-badf-4d54b33deb9d · outbound

This paper cites This implies that in all possible assignments, at least one parent of gi is always 1.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This implies that in all possible assignments, at least one parent of gi is always 1

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.743170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:00.788274Z digest=sha256:edb7cdd0936cad34eec86fdb740123032aa78a5e11bdc44e45c17e029beaa899

Observation a35dfc16-e6ec-4fd6-b17a-0a9593aa35ec · outbound

This paper cites This means that in the collected data, if Xj is one, at least of some other parent Xk is also one.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This means that in the collected data, if Xj is one, at least of some other parent Xk is also one

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.450693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:00.909216Z digest=sha256:c55c0b7316bc51e0a6c377dcf40d0b3da977207b6e5f848fb143bf83802c923e

Observation d5db77dc-54b7-4b3e-bc67-5d6e4c1cdc69 · outbound

This paper cites Thus, the AND operator equals to 0: θi(Xt) = ^ gj ∈PAgi X t j = 0.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Thus, the AND operator equals to 0: θi(Xt) = ^ gj ∈PAgi X t j = 0

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.115544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:01.009013Z digest=sha256:1b1d64fbd89f3b9bc3881503fb3d4eb0313fbe16422bf86587612e371c803055

Observation a059b745-1826-4392-b327-b64a4db82a01 · outbound

This paper cites This means that if Xj is zero, at least of some other parent Xk is one.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This means that if Xj is zero, at least of some other parent Xk is one

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:02.908195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:01.140753Z digest=sha256:0021bfb2b29826329e71256cd5270ca80b12ea767bb78989c979251a789a680b

Observation 5cc87883-af28-4119-a045-0c75be5b4e8f · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 19

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T19:59:02.610404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:01.268863Z digest=sha256:14ff96a7befdaabf869db2357b2afb5ae12325f07b4c54348dd462a8c90e2026

Observation 8419e1d6-20e4-4c15-b4a2-176bb38a79d9 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:02.369210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:01.396496Z digest=sha256:c6103fe06ec15f18a05d189c6c5e91e4ff3f24ed457a57c8133035e68e503944

Observation 6bfd4648-1748-4764-9424-c2d5a7829dae · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 21

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T19:59:02.158132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:01.490528Z digest=sha256:845242b875bd2623febb9ac7867cd36e0b457f7c55f5d25d36cee33285915ce2

Observation c33044bd-59e3-4cb9-aa43-89263fdf337d · outbound

This paper cites oracle goal space.

Hierarchical Reinforcement Learning with Targeted Causal Interventions oracle goal space

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:01.882505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:59:01.602504Z digest=sha256:dc9df2796f3da0e587c4d2b72c0c8b362258036b84ed05cf84bfdac13929ec94

Observation 69459973-4e23-4b37-8620-6369ffb101a3 · outbound

This paper cites MineRL: A Large-Scale Dataset of Minecraft Demonstrations.

Hierarchical Reinforcement Learning with Targeted Causal Interventions MineRL: A Large-Scale Dataset of Minecraft Demonstrations

Reference 1528

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:59.435876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:59.435876Z digest=sha256:bd6be1c354b99cc16e540e6c960ff25fc15af42d604bf199de29b7addeae9381

Observation 9faebd17-940c-4b85-a192-68f2b8d61b45 · outbound

This paper cites Learning Neural Causal Models from Unknown Interventions.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Learning Neural Causal Models from Unknown Interventions

Reference 3119

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:59.513260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:59.513260Z digest=sha256:0465b99ee98c5a8c85e336f99f2d56b4dd3afc7ac2f0b922ec95c317fb48e40e

Pith citing papers

No inbound Pith citation observations are available.