Pith. sign in

Paper Citation Record · LEDGER

Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2403.09472.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.09472 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T20:01:43.731159Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:56:14.059065Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 36cfce41-6f7a-4201-b939-1e7d509b17f3 · inbound

Optimizing Temperature for Language Models with Multi-Sample Inference cites this paper.

Optimizing Temperature for Language Models with Multi-Sample Inference Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T20:01:43.731159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:01:43.731159Z digest=sha256:874d85eafde796a08483ab77b2e68a4a13b01b2a4dd7e592162d67ca1cabacc3

Observation 56a5b2ec-712a-4545-8dd9-8ba54ba819fc · inbound

Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision cites this paper.

Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:20.136880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:14:20.136880Z digest=sha256:12e8340eacae8bc3931bb8c0ee6e596b24cfad4216357df3db16c99680854aa9

Observation 8324132b-df0d-4448-9fec-95ab36f9b18b · inbound

MetaLint: Easy-to-Hard Generalization for Code Linting cites this paper.

MetaLint: Easy-to-Hard Generalization for Code Linting Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T04:12:02.906905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T04:07:31.283348Z digest=sha256:c30d52da46c57a0d1443195994c723966290c52eef9f8a56e5b35502233e0f47

Observation 86982e4b-f595-4084-a8d2-944086c00017 · inbound

CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning cites this paper.

CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:25:45.309813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T23:24:43.556606Z digest=sha256:135f3593edd185ecc62d93c336c0b272df9f892f1bb8a1f299d744f2ba6560fa

Observation 4f6e0567-ee29-4a92-8752-5f461b939a8c · inbound

Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher cites this paper.

Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:56:14.061212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T17:35:24.108984Z digest=sha256:fdb561efc2bc303e3107d0270fc9ab483ee0ea2abd788ba7c1196f99e0ff9782