Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning with Targeted Causal Interventions

As of 9 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2507.04373.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04373 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:59:01.602504Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved8
  • parse uncertain2
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bcd355d3-c603-4b23-a5b9-f8a40174012e · outbound

This paper cites For” loop, it records a trajectory τj. The “While.

Hierarchical Reinforcement Learning with Targeted Causal Interventions For” loop, it records a trajectory τj. The “While

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:06.021733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:58:59.566592Z digest=sha256:42751946fbea4bef8b0006de5eb3f85b7770abb6771f5dac4bd8d04a91cd6654

Observation 95ea9249-855c-411c-a23e-662050c6afee · outbound

This paper cites As we mentioned earlier, q1 is the upper bound on the expected number of the nodes that are ancestors of the node n.

Hierarchical Reinforcement Learning with Targeted Causal Interventions As we mentioned earlier, q1 is the upper bound on the expected number of the nodes that are ancestors of the node n

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.944378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:58:59.618392Z digest=sha256:51517a53d029a460a2d98c91eaa86ba30a2cf10a8cc7046050d2e02bb29908ad

Observation 69954592-9938-4d56-bf2e-2c8dc644aadb · outbound

This paper cites At this point, we have constructed Ijr.

Hierarchical Reinforcement Learning with Targeted Causal Interventions At this point, we have constructed Ijr

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.518464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:58:59.825900Z digest=sha256:eeee02cc33f03ef46fe4fb4aed2fe045d3b0ca38afe1feb70942b96c559930d6

Observation 3c607cc7-9191-47d2-aed3-1feccd0c5f42 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.057351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:00.670188Z digest=sha256:f68190e776302da509f1878001caef8b2d91bf7c5733e1d997b30a1679bbbd65

Observation edde5922-6480-4b98-ab29-786f37d13290 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:05.828947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:58:59.661189Z digest=sha256:ec07c7e89e993d70b6e534e99f0ecb87860ead60185b7558515fcd7e51c4cb5b

Observation 42e29f81-69a3-43a8-b12a-03053e05786f · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:05.693361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:58:59.724253Z digest=sha256:38aa37a8b096c812a3d6c27e3ad591dd564ee07d3f4375308cacfe52f8158fe4

Observation fa4077de-a00f-4075-9185-2e15929efd7f · outbound

This paper cites We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi.

Hierarchical Reinforcement Learning with Targeted Causal Interventions We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.217088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:58:59.911931Z digest=sha256:7f1984b3974493765be877a978498cb33710d222816f7b02abbb2d0c66aec373

Observation 9e8f781a-6a32-4a70-8290-de481f5cb137 · outbound

This paper cites • If gi is the goal, backtrack to construct the path.

Hierarchical Reinforcement Learning with Targeted Causal Interventions • If gi is the goal, backtrack to construct the path

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.993906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:00.025603Z digest=sha256:801b97378aec077590a43927b70ad2aa9d491abc66af2b621a684295cb974e50

Observation 0fdd8088-47d8-400b-8192-c94b8819f64e · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.830033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:00.125185Z digest=sha256:266f77a2ff2522da9108e57d4d10f861024f5e5af505a6c8dae046a894b05218

Observation 1d7564ee-af70-4be3-a534-1afc40a97600 · outbound

This paper cites We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi.

Hierarchical Reinforcement Learning with Targeted Causal Interventions We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.636736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:00.267033Z digest=sha256:e78c85ba8cf0bb06d3d28f097657628318342e12f7597ddcf06bd7b64c0b91fb

Observation 85b7b82b-c5fa-4cc4-a2ce-39d18e44cbd0 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.475273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:00.396920Z digest=sha256:0f4c5e3645984f50bad3ead9b0d1f899a4919a2bf58c8b674b4125c4b616f6fb

Observation 6a16f74e-6055-4dfa-af95-128e97919592 · outbound

This paper cites – For each neighbor gv of gu in ˆGt: * Set w(gu, gv) = |C HˆGt gu \ back track(gu)| + 1.

Hierarchical Reinforcement Learning with Targeted Causal Interventions – For each neighbor gv of gu in ˆGt: * Set w(gu, gv) = |C HˆGt gu \ back track(gu)| + 1

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.279653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:00.547872Z digest=sha256:1f5f3a6d1fe67c7d9414a53b0ed4412fdb90c2ce4c045a0b9878a27ff67272d5

Observation 8eb58f2e-b6e1-4739-badf-4d54b33deb9d · outbound

This paper cites This implies that in all possible assignments, at least one parent of gi is always 1.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This implies that in all possible assignments, at least one parent of gi is always 1

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.743170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:00.788274Z digest=sha256:ece581c0f1c3857999c9f3d6440c80d5d55d0202e153a9082c79f2143ec05138

Observation a35dfc16-e6ec-4fd6-b17a-0a9593aa35ec · outbound

This paper cites This means that in the collected data, if Xj is one, at least of some other parent Xk is also one.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This means that in the collected data, if Xj is one, at least of some other parent Xk is also one

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.450693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:00.909216Z digest=sha256:5d0f95342c5a1ae3f8ea79fbe53830f7f85aab5eaf42f9c96bf12ce3d1030413

Observation d5db77dc-54b7-4b3e-bc67-5d6e4c1cdc69 · outbound

This paper cites Thus, the AND operator equals to 0: θi(Xt) = ^ gj ∈PAgi X t j = 0.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Thus, the AND operator equals to 0: θi(Xt) = ^ gj ∈PAgi X t j = 0

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.115544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:01.009013Z digest=sha256:d2f6181a0f6802c567bd1b98a888a4621e5740a227a43e349be9d37c43e9f6a2

Observation a059b745-1826-4392-b327-b64a4db82a01 · outbound

This paper cites This means that if Xj is zero, at least of some other parent Xk is one.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This means that if Xj is zero, at least of some other parent Xk is one

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:02.908195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:01.140753Z digest=sha256:66b0c8de6eb5f94ec2386e0b231e060e962bc0685ecf14c3339f5ffcb2b15059

Observation 5cc87883-af28-4119-a045-0c75be5b4e8f · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 19

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T19:59:02.610404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:01.268863Z digest=sha256:d5ba8b60e0ddaecc0ceb9d5b3e80cd913b0ce2c80bd01f2ea50f1475780f9af1

Observation 8419e1d6-20e4-4c15-b4a2-176bb38a79d9 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:02.369210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:01.396496Z digest=sha256:f821d80fb83caf44e8960e783dc4ff8aa3c524a93f0a3f14a0552291551a1ace

Observation 6bfd4648-1748-4764-9424-c2d5a7829dae · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 21

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T19:59:02.158132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:01.490528Z digest=sha256:130fb912d58fcb763db7e93976f2905ee5fe0ad0ef2ee9791e304f517f00c15e

Observation c33044bd-59e3-4cb9-aa43-89263fdf337d · outbound

This paper cites oracle goal space.

Hierarchical Reinforcement Learning with Targeted Causal Interventions oracle goal space

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:01.882505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:59:01.602504Z digest=sha256:be5aab033ac3486796e174b41fbec5bbc979be88016ba2db44c3330a336eaab6

Observation 69459973-4e23-4b37-8620-6369ffb101a3 · outbound

This paper cites MineRL: A Large-Scale Dataset of Minecraft Demonstrations.

Hierarchical Reinforcement Learning with Targeted Causal Interventions MineRL: A Large-Scale Dataset of Minecraft Demonstrations

Reference 1528

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:59.435876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:59.435876Z digest=sha256:e959ff2fe66c5bcff2b5c0d7fb6714ce3805737957aa8227edd5d529095563a9

Observation 9faebd17-940c-4b85-a192-68f2b8d61b45 · outbound

This paper cites Learning Neural Causal Models from Unknown Interventions.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Learning Neural Causal Models from Unknown Interventions

Reference 3119

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:59.513260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:59.513260Z digest=sha256:0c4908a2ae392e560db777429254bec374e2b8edf95dc98a0ba700498a3cbeab

Pith citing papers

No inbound Pith citation observations are available.