Pith. sign in

Paper Citation Record · LEDGER

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning

As of 17 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2412.11120.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11120 v2

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:22:54.460782Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T03:10:51.564249Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved9
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 23780dd9-feac-4db3-8e2f-d9fc03c23e38 · outbound

This paper cites an unresolved cited work.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:22:54.622791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.428216Z digest=sha256:00baf33d51d6f8feba9cc7c7c21effb73f53d38c72b14e43a78e110b8fff8e55

Observation 948adef9-2e35-4e7f-8644-097076170bfb · outbound

This paper cites an unresolved cited work.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:22:54.611305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.432704Z digest=sha256:0f8a7ac4c58cdae51db9fe695f5fb53a90c54325d9000cfd789611f98c59e7b1

Observation 3e47ae07-5038-4723-81c5-6e8795e94558 · outbound

This paper cites Agent-Temporal Attention for Reward Redistribution in Episodic Multi-Agent Reinforcement Learning.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Agent-Temporal Attention for Reward Redistribution in Episodic Multi-Agent Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T15:22:54.423523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:22:54.423523Z digest=sha256:704352d16b94276e9b7012b551c2786131c349df8872fbc5254ec18faf6f8858

Observation 7ddfb52d-294b-4112-9578-bde29e46b692 · outbound

This paper cites an unresolved cited work.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Unresolved cited work

Reference 4

Resolution
parse uncertain
raw_fallback, observed 2026-08-11T15:22:54.586856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.440284Z digest=sha256:b791c6b7c4a4df10564679b55510fe46806501d2ca88361541de010cd8ddc4a4

Observation 2791e3d9-a7ea-4007-aceb-eccce2444639 · outbound

This paper cites an unresolved cited work.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:22:54.575215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.444067Z digest=sha256:5c265455e7ab9262c75a4d42a1e1504e19dc841d4a02f1714411e959ac02735a

Observation 73247d3a-6ce9-4f25-a694-22a9b6083a94 · outbound

This paper cites an unresolved cited work.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:22:54.599304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.436494Z digest=sha256:2e50a6557661aad498265615675a8f64e8df7b1f878b1827f44c7a8eecd94b1b

Observation 7552e827-5516-4d57-abd1-d0ea290fe919 · outbound

This paper cites an unresolved cited work.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:22:54.563247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.448113Z digest=sha256:0dfe74b034a52e42160ef3ad31901a540a89b967737c62cba48c12e0c0bebf55

Observation e860c4f1-41e8-4f56-8aeb-081e6e172b52 · outbound

This paper cites an unresolved cited work.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:22:54.551000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.452133Z digest=sha256:6e367b1552cf08b6c2f543d6077ee5357f77a185687511d9d390cb4b1b47f73f

Observation 3a1eca03-65b5-496b-83b7-de341d0f398c · outbound

This paper cites Reward Design.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Reward Design

Reference 11

Resolution
malformed identifier
raw_fallback, observed 2026-08-11T15:22:54.538986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.455885Z digest=sha256:3ec2e50b64118632e34b42e0dd0a09c84f83d9ff48048d43bc3bddd74c941ff6

Observation e3338cc6-963c-46a5-8fbb-1c7401b7bb1e · outbound

This paper cites As shown in Fig.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning As shown in Fig

Reference 500

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:22:54.526899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T15:22:54.460782Z digest=sha256:2848bab2afbbfaf168c2b2d25ef51ab3f51490065d47859cd8fc271e1bd0f435

Observation 27c941b5-3961-43af-ad2f-d36d6cc64edd · outbound

This paper cites Align-RUDDER: Learning From Few Demonstrations by Reward Redistribution.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Align-RUDDER: Learning From Few Demonstrations by Reward Redistribution

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T15:22:54.412698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:22:54.412698Z digest=sha256:9a273e92148784e96b08a1385089fa242a5cefb9ada138bcd13dbb48d6222612

Observation 2f2bbcd7-286a-4a5a-89db-1f91cd5f7619 · outbound

This paper cites Self-Refined Large Language Model as Automated Reward Function Designer for Deep Reinforcement Learning in Robotics.

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning Self-Refined Large Language Model as Automated Reward Function Designer for Deep Reinforcement Learning in Robotics

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T15:22:54.417804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:22:54.417804Z digest=sha256:e6e9ea674bcca0cbc46bec9c2c2fcae73be1127de5f4fa462d6ec186c762dbc7

Pith citing papers

Observation 36014a80-2955-4406-be10-e95963b257ce · inbound

TRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon Agents cites this paper.

TRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon Agents Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T03:10:51.564249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:10:51.564249Z digest=sha256:0614dd6bb815baa34cf891b5606b5075bd1cad7bbaa594727d3d985261a8918b