Pith. sign in

Paper Citation Record · LEDGER

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 1 inbound Pith citation observation for arXiv:2501.17842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.17842 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T04:37:02.239371Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T12:15:08.304150Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T07:59:40.742775Z

Reference resolution

61 of 61 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved23
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a0e42ae-21d7-4791-931a-3b14104fb482 · outbound

This paper cites Critical learning periods in deep networks.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Critical learning periods in deep networks

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.739659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.739659Z digest=sha256:3662ba6954d231cf10d6f84afc13f65a5bf080855683a1d9837a435646c81391

Observation 95c0459c-a4a2-465e-a229-43ca501811ef · outbound

This paper cites Learning dexterous in-hand manipulation.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Learning dexterous in-hand manipulation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.746313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.746313Z digest=sha256:4fbf5bbba2fc0bf10a85b4f7112c62d9fcb61cd6141646b21f60b328d874bb51

Observation 9dd481e3-8c47-4a4f-99de-30193bbc9e40 · outbound

This paper cites A survey on intrinsic motivation in reinforcement learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning A survey on intrinsic motivation in reinforcement learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.752309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.752309Z digest=sha256:7e6b70c04197770a85c7c126de26c0fda74b722269f26f3840c8a63c449aa257

Observation 7e99525c-3df7-406f-896d-d4c77bf09dc6 · outbound

This paper cites Never Give Up: Learning Directed Exploration Strategies.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Never Give Up: Learning Directed Exploration Strategies

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.760052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.760052Z digest=sha256:da19b4b0d3cf9dffe9fae92d9e5e99132b7e54cf882da959f2f0899657fae589

Observation d648ce7f-df74-4c50-a0d5-e3901ce83a42 · outbound

This paper cites Toddler-inspired visual object learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Toddler-inspired visual object learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.620991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.766890Z digest=sha256:84b5b989b743ac5ba1bc04219b41821dc47b3d79ab389e6bc3a50f14693dc7e0

Observation 65f96382-5527-4618-a2c7-8ade189ce4f1 · outbound

This paper cites Curriculum learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Curriculum learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.604620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.774882Z digest=sha256:4a90d27248a1875a52e91cfd965a09063a31f568624973ec5aae610c9579929f

Observation 2566d5af-ecf6-402e-a52d-a80ef319f9b9 · outbound

This paper cites OpenAI Gym.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning OpenAI Gym

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.781050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.781050Z digest=sha256:239ba6020486204b55098ec4ecadfc61abaef8a3368766f206b69d77b4c56372

Observation 95bee900-f074-4859-9416-29d67b4665e0 · outbound

This paper cites Exploration by Random Network Distillation.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Exploration by Random Network Distillation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.787185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.787185Z digest=sha256:e8bb965dd8a9f7c4e610a8c888f145c5dc8caedd77ffd3787f0a6bd88031d4f1

Observation b70eef33-a63e-484f-9bff-a45998d7e0fc · outbound

This paper cites A critical period for robust curriculum-based deep reinforcement learning of sequential action in a robot arm.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning A critical period for robust curriculum-based deep reinforcement learning of sequential action in a robot arm

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.587524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.794058Z digest=sha256:7afbdac107113dc7a3853ae0a8b9cd6dd57ee41a10e2e6643483a446039b5ebe

Observation 915198f8-8bbf-4e04-b03c-201fd8ee75b0 · outbound

This paper cites Class rectification hard mining for imbalanced deep learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Class rectification hard mining for imbalanced deep learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.570077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.799597Z digest=sha256:738f32dcba8642b104c231df28968eae24a89cfa2637c2ee6ecd4acb50abd1aa

Observation f5646f98-1f35-4a8b-bc16-fbacb3d00ff9 · outbound

This paper cites Self-contrastive learning with hard negative sampling for self-supervised point cloud learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Self-contrastive learning with hard negative sampling for self-supervised point cloud learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.805366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.805366Z digest=sha256:8c2cf02bd86e4d4b467fd2d5505222dc3493c1bb7c86b447ad4c5ce9982af7e1

Observation 218cc5a5-0000-49d9-a628-47ec6a2704d2 · outbound

This paper cites Automatic goal generation for reinforcement learning agents.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Automatic goal generation for reinforcement learning agents

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.553506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.810267Z digest=sha256:3acc44204ca0e6c470041b551321a4fb3c0e67a35866c64a1b054e10e462dc4b

Observation 84a96359-289a-4d61-bd5e-3a150cc4a11c · outbound

This paper cites Sharpness-aware minimization for efficiently improving generalization.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Sharpness-aware minimization for efficiently improving generalization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.535508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.816162Z digest=sha256:92ca74357fef6831dec5750489f99707858daaa3f6923fc22a3b40d3ff3f1ce8

Observation 0fc64829-8eb4-457d-a1af-cfd44bd92237 · outbound

This paper cites Exploratory behavior in the development of perceiving, acting, and the acquiring of knowledge.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Exploratory behavior in the development of perceiving, acting, and the acquiring of knowledge

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.518317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.821217Z digest=sha256:34183c5498e6fde93114478933afaa0f044d17b0837f91acd920fe417570d6d9

Observation 0e2f4b8c-b6f7-4732-bd8e-896e135c38fc · outbound

This paper cites Qualitatively characterizing neural network optimization problems.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Qualitatively characterizing neural network optimization problems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.828345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.828345Z digest=sha256:521784ca014212c813a7455ddf68d125f58606ec0d35548b23da49fad571cb22

Observation 3d3c6b86-58c7-4989-bb18-16d15ba4cec7 · outbound

This paper cites The scientist in the crib: Minds, brains, and how children learn.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning The scientist in the crib: Minds, brains, and how children learn

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.500009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.834222Z digest=sha256:27fd094858274773d36e68cb0f6629f8d06675033ee220946cb5b6464138b03d

Observation 55b368b1-6db1-4525-b3fb-5de18650d1cd · outbound

This paper cites Changes in cognitive flexibility and hypothesis search across human life history from childhood to adolescence to adulthood.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Changes in cognitive flexibility and hypothesis search across human life history from childhood to adolescence to adulthood

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.482204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.839444Z digest=sha256:7cb03350391acab75f07824339c3e7ae1313aa64ea85d749dcc7be5a5418c397

Observation 202da60a-af40-4e8b-af47-ef38ccc6c3e3 · outbound

This paper cites Automated curriculum learning for neural networks.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Automated curriculum learning for neural networks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.465050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.844547Z digest=sha256:ef171313bdd65dfc7a78182688a8f19a07a2301b79981367addc272d4f2242fa

Observation 0061575f-658f-4225-b700-bc81de7be184 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.849912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.849912Z digest=sha256:1e9647ac250f3c6f90c916b3403bc40a5fe620e5c3d282ce947dfe877f6c8f62

Observation c09b2fce-b3d0-4755-809b-1c130020064a · outbound

This paper cites On the power of curriculum learning in training deep networks.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning On the power of curriculum learning in training deep networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.431908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.855034Z digest=sha256:e91d0fe454241e380fdba4baa3311600875a3b12a872bf633b09a9d800fd6a5a

Observation bfb38035-4bbb-48cc-85b1-f078aa19bae0 · outbound

This paper cites Dealing with Sparse Rewards in Reinforcement Learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Dealing with Sparse Rewards in Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.861960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.861960Z digest=sha256:b40b9987f269d2e3115787f609e4d6fb62e13eb46fe5ceb55c2de09413a54ca3

Observation 3221bf7a-5b74-4e77-952a-b02bca32e2e2 · outbound

This paper cites Expressing arbitrary reward functions as potential-based advice.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Expressing arbitrary reward functions as potential-based advice

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.414732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.869271Z digest=sha256:710dd21d9047d6f77cc6d668aa0b95a8ae29d1325ad9063e62d96e35e41598c2

Observation de85f9db-3569-4201-81f3-3c6b661d49ff · outbound

This paper cites Comprehensive overview of reward engineering and shaping in advancing reinforcement learning applications.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Comprehensive overview of reward engineering and shaping in advancing reinforcement learning applications

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.397635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.875305Z digest=sha256:211e00402539f9b0e3a1ea7468790d0b0e4f7052971483b7e4986ec23c4bce0d

Observation 55b4ce60-6fee-4916-8798-9866b732270a · outbound

This paper cites Finding flatter minima with sgd, 2018.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Finding flatter minima with sgd, 2018

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.377585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.882211Z digest=sha256:e90b6162c69908e53df1349d46f64510751eab466af688d38c2440282fb703d4

Observation 329ac9c2-6788-471c-b331-e90037ecfeba · outbound

This paper cites Hard negative mixing for contrastive learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Hard negative mixing for contrastive learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.360854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.889915Z digest=sha256:1096e8483423ce459041436a778b833ffa0f39724c127969e87c00baf4d989bc

Observation 7e9585ef-5542-430e-b4d2-cf52bfbda266 · outbound

This paper cites Vizdoom: A doom-based ai research platform for visual reinforcement learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Vizdoom: A doom-based ai research platform for visual reinforcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.894875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.894875Z digest=sha256:f51f6c6a6721c4a22cbd6f2a7492d38fb6ffd91f95188f89e5dd8abd0f99a9cc

Observation 2c376e58-dea9-4c00-a18c-2798bea7406e · outbound

This paper cites On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.899981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.899981Z digest=sha256:b32210e505a2595ba7ba5ffe87551b4a4fca2638ee0bff58d6c3fbb6d83e88ad

Observation 2339c1f9-fe19-469e-b2d4-7ba67b75e7c6 · outbound

This paper cites On large-batch training for deep learning: Generalization gap and sharp minima.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning On large-batch training for deep learning: Generalization gap and sharp minima

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.324828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.905558Z digest=sha256:c32fd58569f93d4538770211593963750340e2014141062e9bc78e00dbc4469a

Observation 6c5cbfdf-b4cf-4327-a519-d0155edcd1ae · outbound

This paper cites Goal-aware cross-entropy for multi-target reinforcement learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Goal-aware cross-entropy for multi-target reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.280366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.916395Z digest=sha256:b4bde65f22fd733fb61a5b0f2296c8b3398fe3732aaceb2ae31e2afb523ea97a

Observation 04386915-3dd0-459a-a63e-06dce0ffa9de · outbound

This paper cites L-SA: Learning Under-Explored Targets in Multi-Target Reinforcement Learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning L-SA: Learning Under-Explored Targets in Multi-Target Reinforcement Learning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-10T04:37:02.435838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.921280Z digest=sha256:9147437856959e8c64984475aa87a5579d07195ddb2b7d71b35d70f99d973d5f

Observation 6cbd135b-887d-46d3-ab64-5bd6289b57b4 · outbound

This paper cites Visual Hindsight Self-Imitation Learning for Interactive Navigation.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Visual Hindsight Self-Imitation Learning for Interactive Navigation

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-10T04:37:02.407567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.926810Z digest=sha256:b1a996fa166f456a6f4dd9d76483c937c2dede921d507e4b8f6da97821b76f29

Observation 9cfc33a3-154e-4075-a9fe-c0d5fbeddf36 · outbound

This paper cites Reward (mis) design for autonomous driving.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Reward (mis) design for autonomous driving

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.933335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.933335Z digest=sha256:e2d28bd67b1d50f4b25fe05d06dd3caeb71a42b80be145d726a70277f6f1c606

Observation 1a1fb464-a246-4375-aec4-0f20a0a8cf67 · outbound

This paper cites Exploration in deep reinforcement learning: A survey.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Exploration in deep reinforcement learning: A survey

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.244945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.939931Z digest=sha256:a7ed8fe4f729d7c917182fc1624430c2eb004ddf1e47550d90051ed210fe38a2

Observation ed9c321e-01e8-4573-a784-069762bf7b50 · outbound

This paper cites Theory and application of reward shaping in reinforcement learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Theory and application of reward shaping in reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.217213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.945596Z digest=sha256:e4fc43080f5be2d10815c56fa18f8284cad5e51ae75204bb9acd37457622dc9a

Observation e0839428-7a16-4cde-9b54-4ad5c0550fb4 · outbound

This paper cites Visualizing the loss landscape of neural nets.Advances in neural information processing systems, 31, 2018.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Visualizing the loss landscape of neural nets.Advances in neural information processing systems, 31, 2018

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.951105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.951105Z digest=sha256:a480d0a5383bf75964bdee4eb82a76ee2e5df71fa6f6209500c23638739519a7

Observation 87f954d8-2606-4997-b33a-cc02cef768a4 · outbound

This paper cites Continual reinforcement learning in 3d non- stationary environments.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Continual reinforcement learning in 3d non- stationary environments

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.173903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.962869Z digest=sha256:9f3102707d8f1a9c235c86d921b222748e2972fe63780015357b81ad81d05c5e

Observation f65ea873-5742-4f77-9b01-ab806b08bc78 · outbound

This paper cites Information-based objective functions for active data selection.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Information-based objective functions for active data selection

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.153141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.968890Z digest=sha256:0a5638708b4cab99520eab03e4028eed1841045e640c8fec2d5518ff40f6f409

Observation 1feb27b2-ee35-4b99-8dd9-d4e4be6d5500 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.976047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.976047Z digest=sha256:57b8e971eaf5585f93923377f59d6926dd3d468556d14afff4f9676acfecbf2d

Observation 5d951831-c388-49d0-ac77-eda90a41b493 · outbound

This paper cites Rusu, Joel Veness, Marc G.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Rusu, Joel Veness, Marc G

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:01.981923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:01.981923Z digest=sha256:1cb2a48d19808f138d5d9d5368e0fec49ad36c02d258259e031266567b8b0f60

Observation 72f4ae6d-b28d-47a3-a670-c8af144f8283 · outbound

This paper cites Asynchronous methods for deep reinforcement learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Asynchronous methods for deep reinforcement learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.135109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.988468Z digest=sha256:c605111e76e42582ca69624cbb6be7d5b5a3f42faecd207dae15e1a1c85466ac

Observation c673ddb6-2ad3-4664-b345-4f759ac75097 · outbound

This paper cites Generalizing curricula for reinforcement learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Generalizing curricula for reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.117560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.996467Z digest=sha256:b5ef3d36d6cf7630b3118d606b5d09f4c64b4b0b4100055ecf1a3bdb33d094cc

Observation 94bb4d39-cc78-4fda-aac4-107954a11654 · outbound

This paper cites Ng, Daishi Harada, and Stuart J.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Ng, Daishi Harada, and Stuart J

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.101108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.001571Z digest=sha256:cbf5916a74d3925d698bcd5855d6f5968159dee74d20ae31f2be1116bf2c6523

Observation 377bf4d5-35aa-4735-a371-6ae1c7992d08 · outbound

This paper cites Policy invariance under reward transformations: Theory and application to reward shaping.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Policy invariance under reward transformations: Theory and application to reward shaping

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.080041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.006759Z digest=sha256:e0b97dff0c905c95d389af9b48bea22da3bf069bd3622269f4043b8aa9e27f10

Observation 9750cda7-d33f-4743-aadc-63646e131db2 · outbound

This paper cites How evolution may work through curiosity-driven developmental process.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning How evolution may work through curiosity-driven developmental process

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.061401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.011672Z digest=sha256:2b7ed939b622f9721682803b3dab68a0493ca8e24f4f62745919cd87a61591d1

Observation 06b8e9db-e691-4cd8-8c09-5e9a7dd77f54 · outbound

This paper cites Benchmarking multi-agent deep reinforce- ment learning algorithms in cooperative tasks.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Benchmarking multi-agent deep reinforce- ment learning algorithms in cooperative tasks

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.037321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.016440Z digest=sha256:5b508326a469cdda57b66c6d01d4b360556a249166801b2b279e801ca26c7c61

Observation c6871958-9ad0-4323-81bd-43e01ea9fd19 · outbound

This paper cites Toddler- guidance learning: Impacts of critical period on multimodal ai agents.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Toddler- guidance learning: Impacts of critical period on multimodal ai agents

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:03.017617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.021482Z digest=sha256:5b57707441a11f91770145ac2214e37c979555b8bb9d8034fa5eece2155ef8ab

Observation 6077acab-c7df-4869-9b9f-d37645447845 · outbound

This paper cites Unveiling the significance of toddler-inspired reward transition in goal-oriented reinforcement learning.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Unveiling the significance of toddler-inspired reward transition in goal-oriented reinforcement learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:02.998943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.026176Z digest=sha256:5dacd9c3e4cd00e158973756a99f708fbddc3342eec2f283e33d68e61aa8ca6d

Observation b7070522-7609-4216-8302-a2d98747092a · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Curiosity-driven exploration by self-supervised prediction

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:02.162181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:02.162181Z digest=sha256:878df14b44d93197fb683c9bfe5360d629a90f8110cf20cca4acbb824104b486

Observation c2ffe56c-122d-4b58-b0a0-010e640df1e3 · outbound

This paper cites The origins of intelligence in children, volume 8.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning The origins of intelligence in children, volume 8

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:02.967628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.169483Z digest=sha256:a875886960c028f5815200e4b15925bd56da6aa702a2e380122000390708ccbb

Observation b6ff1b5d-d3a6-4afb-aa5a-3a54ee8fe99d · outbound

This paper cites Stable-baselines3: Reliable reinforcement learning implementations.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Stable-baselines3: Reliable reinforcement learning implementations

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:02.949208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.174900Z digest=sha256:76f2f1ce5a792b3114626b1c0f8627944ef8f8e0aa9eb76a2ee789e38314be59

Observation 5084a7dd-85d9-43fc-9f72-c400a3afb3eb · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:02.180292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:02.180292Z digest=sha256:d2df2a63d4da316bd499a5a9b491ce12c60692b2f1b401474c151d3cd98fdd18

Observation 31e1cd85-c28b-42c0-ba50-c01527ca47b1 · outbound

This paper cites Proximal Policy Optimization Algorithms.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:02.188868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:02.188868Z digest=sha256:4f812bfa480abaa879d2ae12529d3c0a2ac805029cc6ef09fc46ad5142d399e2

Observation 235e6a0c-5ad7-4fb6-ab1d-7ed47b1f8a01 · outbound

This paper cites From neurons to neighborhoods: The science of early childhood development.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning From neurons to neighborhoods: The science of early childhood development

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:02.927361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.194726Z digest=sha256:31482049db4aebfe45681b52c601ce88c5e6dddc0251b4479b2e466d4fcce942

Observation ff5afa0c-e470-492c-8ab2-95fc416d62e6 · outbound

This paper cites Transfer learning for reinforcement learning domains: A survey.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Transfer learning for reinforcement learning domains: A survey

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:02.901494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.201237Z digest=sha256:c1370ac44befdb6e1d90b9d5e4ccbf384265c5ed961b73e9b94d1459d391a98f

Observation b009c179-2eed-4b14-9a0c-4ac844f40e5a · outbound

This paper cites Mujoco: A physics engine for model-based control.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Mujoco: A physics engine for model-based control

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:02.207964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:02.207964Z digest=sha256:c5d61b7df1d32a883aca0b4bbf00787fd77d15e329e0de1ab9cfaa12edb3e36b

Observation 77112c7c-156b-40d2-a2ee-96e7b1ea909d · outbound

This paper cites Cognitive maps in rats and men.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Cognitive maps in rats and men

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:02.213320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:02.213320Z digest=sha256:12fdd8eec93eebd84cde88f74b84b67acbaae79a8e27f7724b2258df751ec33e

Observation ef495413-98ac-48f6-936f-c606bff9f49b · outbound

This paper cites Safe reinforcement learning via curriculum induction.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Safe reinforcement learning via curriculum induction

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:02.857730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.218907Z digest=sha256:cb2b46b33c4de5f3ffa64db3bc1995ebd12d285597b75566f2483118ca43f34d

Observation bfd574d5-3452-412d-88c0-0fa1215993a4 · outbound

This paper cites Curriculum learning by transfer learning: Theory and experiments with deep networks.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Curriculum learning by transfer learning: Theory and experiments with deep networks

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T04:37:02.830977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.224897Z digest=sha256:bd5a7a90d7e4c9fddfc97961109bf343d357cc032b6dd81c68444148e89f5e50

Observation 156d6535-e62d-4ad4-b6ed-8181b30129fc · outbound

This paper cites FRESH: Interactive Reward Shaping in High-Dimensional State Spaces using Human Feedback.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning FRESH: Interactive Reward Shaping in High-Dimensional State Spaces using Human Feedback

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:02.231133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:02.231133Z digest=sha256:6e9ad2ac0473dea226db6d691def9233feb3c22baf8b3a51af17961705b1e769

Observation f1382a11-bbd9-438a-ae39-12fa51c4fe03 · outbound

This paper cites requested.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning requested

Reference 60

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T04:37:02.808327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:02.239371Z digest=sha256:439e367db6236250fcf815bfbc7fd009ee6adf3b91c4cc17df6a78ae7aa424a0

Observation c24ef4d4-4270-44ea-8186-0e0c11f39a48 · outbound

This paper cites an unresolved cited work.

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning Unresolved cited work

Reference 2017

Resolution
unresolved
raw_fallback, observed 2026-08-10T04:37:03.305336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:37:01.911304Z digest=sha256:222d6595aaea534a47c1ae2a1ab22b4f253a021216b2428158ad9c18953473a4

Pith citing papers

Observation bad41099-7891-40e3-b218-6fc159836bcf · inbound

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning cites this paper.

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning

Reference 144

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:59:40.744288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T12:15:08.304150Z digest=sha256:6293123015722852993cc3b45d0e8d38724f778f5c614e965bfce482fbc015dd