Pith. sign in

Paper Citation Record · LEDGER

Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2408.03029.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.03029 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:26:09.307525Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:44.620366Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fc467a6f-ffc3-44be-8255-83d1f2edd583 · inbound

Skill Expansion and Composition in Parameter Space cites this paper.

Skill Expansion and Composition in Parameter Space Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:09.307525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:09.307525Z digest=sha256:effb0ba20bbcc897cf08c7568fb009c91f9f15e8cb83a3a49439c8c8bdbe5546

Observation 3e597636-2626-49ea-8caf-c9ffcbacd396 · inbound

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning cites this paper.

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:24:47.176236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T00:19:48.053466Z digest=sha256:f82e77b7f65e314fcf3b17c37a619a244384d0897812bd0a8f55b8642c805fd9

Observation 5d24a28d-6e8f-403d-8d68-7a6b57df9834 · inbound

Learning Process Rewards via Success Visitation Matching for Efficient RL cites this paper.

Learning Process Rewards via Success Visitation Matching for Efficient RL Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:59:44.621838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T09:20:35.062060Z digest=sha256:1d75c8f7c47bec204c2c2be99ceccf64afd539dd907821f1ed455bc895b7b27c

Observation 6d0d0be8-2b84-40f4-a650-31a938adcc53 · inbound

Information-Based Exploration via Random Features for Reinforcement Learning cites this paper.

Information-Based Exploration via Random Features for Reinforcement Learning Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T16:35:01.478900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T16:35:01.478900Z digest=sha256:693de55a99d9cd10075bc08129fcd81566110e96a870c21651fd858e4cb8c4d2

Observation 60712595-d350-4103-9b40-62289167d629 · inbound

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback cites this paper.

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T03:14:11.955077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:14:11.955077Z digest=sha256:b037f0cda094d20809fbaca8fe577fbfc048fefb06fc892d75afedfb5420d0e4