Pith. sign in

Paper Citation Record · LEDGER

Meta-learning how to Share Credit among Macro-Actions

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.13690.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13690 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:33:25.016687Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact2
  • verified fuzzy18
  • unresolved17
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c032ebbc-9c2e-46ee-a361-fe1d88d5b00b · outbound

This paper cites Mas- tering the game of go with deep neural networks and tree search.

Meta-learning how to Share Credit among Macro-Actions Mas- tering the game of go with deep neural networks and tree search

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.784818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.784818Z digest=sha256:5822ad9d7191ef290df285fa8750d2695c77947ac7d757f886cb3ff99fc8dfca

Observation 5293bb1b-0cb6-4997-8349-c0e0c202749c · outbound

This paper cites Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H.

Meta-learning how to Share Credit among Macro-Actions Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.815439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.815439Z digest=sha256:ebe1ad5ae82c111dcc9147930c2e1049b5a65d0605390b0433972cedb8a557be

Observation efaa988e-376d-495e-b48f-edcb56b3070a · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Meta-learning how to Share Credit among Macro-Actions Dota 2 with Large Scale Deep Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.820678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.820678Z digest=sha256:0a5783877986a2c905fec9beac3997d6b74c3fd87e363a069fd509b3d0f3d0a6

Observation 1550dea7-2e6d-4b8a-8c3e-723b6dcf7f3e · outbound

This paper cites Autonomous navigation of stratospheric balloons using reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Autonomous navigation of stratospheric balloons using reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:28.860074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.824827Z digest=sha256:341da48d6a081c11c349daebd5de2c6aef53b2b0efa0c2348eb6b3602526baf1

Observation e134d38b-7f0c-48d9-8251-0e86ca64c058 · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Magnetic control of tokamak plasmas through deep reinforcement learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:28.618915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.834321Z digest=sha256:c1eb319959e787c820a3351845c2cc50cc674c16946a3dc144c414db5e3bc5f7

Observation b97099db-f115-465c-869a-d62877af7863 · outbound

This paper cites an unresolved cited work.

Meta-learning how to Share Credit among Macro-Actions Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:33:28.416804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.838471Z digest=sha256:0d9cc2ccb00acfa8358e8bbbe8b05e25ec1a90f2bc93546c7cff8cfd4d1b9c8b

Observation 89f20de9-1747-46ab-b6b9-641b87c887f6 · outbound

This paper cites Hierarchical solution of markov decision processes using macro-actions.

Meta-learning how to Share Credit among Macro-Actions Hierarchical solution of markov decision processes using macro-actions

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:28.120610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.847801Z digest=sha256:bb5cd4b9898a10a3a9cd742c4a71ef9946a8443e498b2915097166e7860def08

Observation d911cf51-915f-40ee-a8f0-d1eec8604439 · outbound

This paper cites Fikes and Nils J.

Meta-learning how to Share Credit among Macro-Actions Fikes and Nils J

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.851769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.851769Z digest=sha256:4db58de6487df863cc9e07fba1bc7bec1d920f2e57d766a037bcd5cb31c614ef

Observation 6d934cc5-c425-4162-9b9f-d21b743f5ae3 · outbound

This paper cites an unresolved cited work.

Meta-learning how to Share Credit among Macro-Actions Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:33:28.020322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.855868Z digest=sha256:5a4e12e4707483573adb839a05d6ad5201411e52b869416ed4b334085719be8c

Observation 10a3d0ce-5480-461f-abac-770f3db7d5b3 · outbound

This paper cites Durugkar, Clemens Rosenbaum, Stefan Dernbach, and Sridhar Mahadevan.

Meta-learning how to Share Credit among Macro-Actions Durugkar, Clemens Rosenbaum, Stefan Dernbach, and Sridhar Mahadevan

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:27.845972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.859868Z digest=sha256:e7b5045b5c4e411231ffcc0d975d31e4f8badfa996c38441e581a2fcda726f17

Observation bd294195-f6c9-4baa-9985-dd1a27100afa · outbound

This paper cites Rainbow: Combining improve- ments in deep reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Rainbow: Combining improve- ments in deep reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:27.642436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.863690Z digest=sha256:e5991492ae92f7cfcac1f83433342c230c0c92dc02255f93ed4e194463563d5b

Observation 43c626f0-e847-4ef6-b22b-87cd80632f26 · outbound

This paper cites Learning macro-actions in reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Learning macro-actions in reinforcement learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:27.346950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.867588Z digest=sha256:9d42ee299aff6706b3332b355613c54b0718c118e0dedf01a0e0c92a64196864

Observation f83086f3-342b-4f35-96e1-961be60390b2 · outbound

This paper cites Macro-actions in reinforcement learning: An empirical analysis.

Meta-learning how to Share Credit among Macro-Actions Macro-actions in reinforcement learning: An empirical analysis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:27.134155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.871801Z digest=sha256:5b8464aeb62a40fc015cf888ecce757961b618635b00f234bd91ada9f7da447c

Observation aff8175a-875f-48c2-adaf-321be739ae24 · outbound

This paper cites Meta learning shared hierarchies.

Meta-learning how to Share Credit among Macro-Actions Meta learning shared hierarchies

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.896757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.876353Z digest=sha256:46f6dd02b165dd130298ceb75bf4a6bd0fbe3d4cfdd567f5dab248e1b41604bf

Observation e1bf3579-d282-4194-b486-5fac2345c4ce · outbound

This paper cites Hierarchical Meta-Reinforcement Learning via Automated Macro-Action Discovery.

Meta-learning how to Share Credit among Macro-Actions Hierarchical Meta-Reinforcement Learning via Automated Macro-Action Discovery

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.881222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.881222Z digest=sha256:824249d35d8dc9a53eb2234f4fbd2dad00863e45d8ec160e0df24b7eb289acc4

Observation dacd4186-125f-479b-aaf4-cf4c4eaf1a89 · outbound

This paper cites Deep reinforcement learning for decentralized multi-robot exploration with macro actions.

Meta-learning how to Share Credit among Macro-Actions Deep reinforcement learning for decentralized multi-robot exploration with macro actions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.767158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.885488Z digest=sha256:2c7a166d6db31803a75c3bc8f53de638c1aa81863098ad3ffd3cac93185c6fdc

Observation 70379299-e73c-4cc0-9a69-730e3e2b758f · outbound

This paper cites Macro-Action-Based Multi-Agent/Robot Deep Reinforcement Learning under Partial Observability.

Meta-learning how to Share Credit among Macro-Actions Macro-Action-Based Multi-Agent/Robot Deep Reinforcement Learning under Partial Observability

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.537499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.890230Z digest=sha256:65c9fcf62537aa24763495ed6c1b0f6d2ec2e5d867153be2569783878056ac7c

Observation cd121df8-6703-47da-97dd-62af7a387c5c · outbound

This paper cites Unlocking new strategies: Intrinsic exploration for evolving macro and micro actions.

Meta-learning how to Share Credit among Macro-Actions Unlocking new strategies: Intrinsic exploration for evolving macro and micro actions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.359326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.893857Z digest=sha256:0292e3aa57502bbd5a86b0e0b54427a26f3f6fa7a47132a4800c560204258ca7

Observation 08dc6753-272d-4a80-9ee9-df5b85750ee3 · outbound

This paper cites Reusability and Transferability of Macro Actions for Reinforcement Learning.

Meta-learning how to Share Credit among Macro-Actions Reusability and Transferability of Macro Actions for Reinforcement Learning

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:33:25.228396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.897577Z digest=sha256:6924c7ad6791be5085c3f971255573a03b5e89d79396e2e1182886d0d1dce753

Observation ebba5e75-b913-477f-9a69-37e3efbe2511 · outbound

This paper cites Efficient Black-Box Planning Using Macro-Actions with Focused Effects.

Meta-learning how to Share Credit among Macro-Actions Efficient Black-Box Planning Using Macro-Actions with Focused Effects

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:33:25.197756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.901782Z digest=sha256:1fb12aa43c24b112b988b653cc7a824231de1c32b6085598e97ab86ce7b07d88

Observation fa82daaf-83d0-41d2-a3f4-192602222500 · outbound

This paper cites Learning macro-actions for arbitrary planners and domains.

Meta-learning how to Share Credit among Macro-Actions Learning macro-actions for arbitrary planners and domains

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.275840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.906030Z digest=sha256:71df7b3a58241fb481fb0b6cc280e7c1f45e9127c16aaaf6cba8d55e8e8350d3

Observation e5866fb1-179f-47e0-9711-492435eb9a86 · outbound

This paper cites Modeling and planning with macro-actions in decentralized pomdps.

Meta-learning how to Share Credit among Macro-Actions Modeling and planning with macro-actions in decentralized pomdps

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.202469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.909481Z digest=sha256:3af3304769d8b21eaef4a91310d639a3cf6ab98c942f30837b5a173c316c73f9

Observation 33e48015-0851-484f-b36e-d3619ed0a3b5 · outbound

This paper cites MAGIC: Learning Macro-Actions for Online POMDP Planning.

Meta-learning how to Share Credit among Macro-Actions MAGIC: Learning Macro-Actions for Online POMDP Planning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.917700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.917700Z digest=sha256:443443bd11e3cc0b1e530c78f97c0176af138fbe15b777da6629437ef3f2c65e

Observation d8b605d1-db0b-4be6-8d9c-0fecf3c76f89 · outbound

This paper cites Deep Reinforcement Learning Based Navigation with Macro Actions and Topological Maps.

Meta-learning how to Share Credit among Macro-Actions Deep Reinforcement Learning Based Navigation with Macro Actions and Topological Maps

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.924438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.924438Z digest=sha256:c5d6a28b3b73ae12a0654b94a212ce2edebb1a2b6d6b3bb594eb07cf0e42b5e6

Observation 7b78c45f-3d47-4264-aa80-9701fb320388 · outbound

This paper cites Human-level control through deep reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Human-level control through deep reinforcement learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.928615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.928615Z digest=sha256:ebc7c8ee0a411926ddf3250a3fe233e40c80dea44bf054393537212309183cfc

Observation 2dd9f7fc-0919-41ec-8840-80b3e8c53152 · outbound

This paper cites Bellemare, Will Dabney, and Rémi Munos.

Meta-learning how to Share Credit among Macro-Actions Bellemare, Will Dabney, and Rémi Munos

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.074752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.932262Z digest=sha256:4be958af5dcc40bc149f08ef99279be0051ca45fe23458bacf625ec85d956d4b

Observation 3938a167-c961-4f21-bd21-b3c6b175e0e0 · outbound

This paper cites an unresolved cited work.

Meta-learning how to Share Credit among Macro-Actions Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:33:25.914748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.936341Z digest=sha256:a74ef066ceb029266e2d80235de92afeba6a9b241834350d60fe3ed466fcfaa3

Observation e52ef023-d402-4c67-9239-969339b33889 · outbound

This paper cites Prioritized Experience Replay.

Meta-learning how to Share Credit among Macro-Actions Prioritized Experience Replay

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.940102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.940102Z digest=sha256:7ee5074a03c7216b2a7c8ae04fa63273f64db7580cd5286479b1effc43430ba8

Observation 23c92e8c-3033-433a-b610-33cdbee5ce27 · outbound

This paper cites Deep reinforcement learning with double q-learning.

Meta-learning how to Share Credit among Macro-Actions Deep reinforcement learning with double q-learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.944182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.944182Z digest=sha256:4e91750e6ea1bbcb1841497ceb7bbf3b664d95a0209b5190e4d8e35338c88959

Observation 32133ac0-1f1e-44bb-ac77-14e07dfc6df8 · outbound

This paper cites Dueling network architectures for deep reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Dueling network architectures for deep reinforcement learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.948639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.948639Z digest=sha256:5afef57cab0bedf24a6fc1f1710e6f8e8fa179ed359414c12f45d2504a88cd1f

Observation da0d07e9-62a7-43c6-8d8d-cbae40b1b60c · outbound

This paper cites Noisy Networks for Exploration.

Meta-learning how to Share Credit among Macro-Actions Noisy Networks for Exploration

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.952368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.952368Z digest=sha256:6a0bc8625f803f1c2a85eb2cb6620880590a675f271ea3a1433cd74373c0a796

Observation 783e554d-9d9f-441d-8ac5-76e368f69599 · outbound

This paper cites Meta-gradient reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Meta-gradient reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:25.665399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.956996Z digest=sha256:d5881ed5d66e3f3920b7e8e16657d965499e8c1b7be27d338681665c2a996a10

Observation 73968742-7555-49a2-9de1-b2ba74f34a83 · outbound

This paper cites Universal value function approxima- tors.

Meta-learning how to Share Credit among Macro-Actions Universal value function approxima- tors

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:25.559712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.967341Z digest=sha256:3f5f10d190905f5d04639541f3fd675b1218fc363fbcd2fc97fb2a0d61f975d1

Observation 288709fa-0c4b-49f7-b2d3-d811e4c5cf4f · outbound

This paper cites The arcade learning environment: An evaluation platform for general agents.

Meta-learning how to Share Credit among Macro-Actions The arcade learning environment: An evaluation platform for general agents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.972757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.972757Z digest=sha256:bf731a25dff9426b78cbf486fed3a3a9829397c0274c86a60aa30b0443016d58

Observation a756d7de-1916-4c1b-a497-9b7424264729 · outbound

This paper cites OpenAI Gym.

Meta-learning how to Share Credit among Macro-Actions OpenAI Gym

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.980907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.980907Z digest=sha256:2c67cef36fd5b022ed6364cc4f6b6c43d55781517171079b0227a9682166a417

Observation 4a619404-64c9-49f3-990e-b2e22582241b · outbound

This paper cites The Atari Grand Challenge Dataset.

Meta-learning how to Share Credit among Macro-Actions The Atari Grand Challenge Dataset

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:33:25.057740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:24.996820Z digest=sha256:712c5777795eaa8c32b889674272a5e15abf2b896833fde6ef50ff888aa294f8

Observation 75bbc518-cc62-4693-b51f-183392014552 · outbound

This paper cites Gym-minigrid: Minimalistic gridworld environment for openai gym.

Meta-learning how to Share Credit among Macro-Actions Gym-minigrid: Minimalistic gridworld environment for openai gym

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:25.504282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:25.001724Z digest=sha256:ea998fec05b61d6b6f787cb9edb2a0f95e25c4f84e61cba73d103fc9db70083f

Observation bfe23e56-c531-443b-a03e-b43f469d379a · outbound

This paper cites - Perform a standard TD update with the MASP penalty, updating θ → θ′ using Σ fixed.

Meta-learning how to Share Credit among Macro-Actions - Perform a standard TD update with the MASP penalty, updating θ → θ′ using Σ fixed

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:25.489509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:25.005793Z digest=sha256:9cb46b160ddb50121d344d6cf24485d59e3cce7c63a42ed1017a0e0ca2f4bd8f

Observation e3f33879-12df-4e55-8281-c15b9750d088 · outbound

This paper cites - Evaluate the performance of the updated θ′ using a meta-objective (the standard TD loss).

Meta-learning how to Share Credit among Macro-Actions - Evaluate the performance of the updated θ′ using a meta-objective (the standard TD loss)

Reference 39

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:33:25.449892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:33:25.016687Z digest=sha256:33f576747d81cd853aa32f0e2075e7eb9c3164f3021250ef099a68bbd946d9c7

Pith citing papers

No inbound Pith citation observations are available.