Pith. sign in

Paper Citation Record · LEDGER

Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2603.19453.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.19453 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:28:29.274754Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-01T13:55:46.656662Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation aa1bb373-703b-47d6-9f35-82a72d3f36f1 · inbound

Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon cites this paper.

Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-02T04:04:30.068390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-12T04:11:08.427195Z digest=sha256:36216eac3fef9f5da95dbf06ddab1968d16a8232e5e96f1e87cd002e4ba250b3

Observation d73cec6b-3b1e-45d5-b488-7ddf8df11e60 · inbound

Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon cites this paper.

Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T13:55:46.658526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T22:28:56.528033Z digest=sha256:ed3ffeca86697071eb967ca86d294f4fa5675dbaf89599b7d542e81d71f6051d

Observation 513985fb-12da-413a-b227-501a0b526bf2 · inbound

Discovering Cooperative Pipelines: Autoresearch for Sequential Social Dilemmas cites this paper.

Discovering Cooperative Pipelines: Autoresearch for Sequential Social Dilemmas Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T00:12:50.014225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T00:08:40.245724Z digest=sha256:214b47ce9b3718fd073b864cacab59f434207584e8d442c5eb9dc478eb1ad48e

Observation f51176cd-a51d-4833-92ac-0f9649304b43 · inbound

Training Small LLMs as Spatial Multi-Agent Policies cites this paper.

Training Small LLMs as Spatial Multi-Agent Policies Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.064587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.064587Z digest=sha256:9059629bbc3fdf7e8d51e67340fcee005eaa47ae67b82437f91dd91885efb770

Observation 26234f6a-9359-4f30-9fd3-9948ac789979 · inbound

A Hybrid Nested Harness for Decoupling Structure and Parameters in LLM-Driven Optimization cites this paper.

A Hybrid Nested Harness for Decoupling Structure and Parameters in LLM-Driven Optimization Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-12T00:28:29.274754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:28:29.274754Z digest=sha256:2389519547c2cfadbfa00d2813d7fc0b58e080f3713c7f18b251ba7d7c7b5a44