Pith. sign in

Paper Citation Record · LEDGER

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback

As of 11 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2507.21131.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21131 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:14:16.727994Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 98fea892-9983-4f3c-bf58-55ae314d04c5 · outbound

This paper cites Concrete Problems in AI Safety.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Concrete Problems in AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.642328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.642328Z digest=sha256:a9209ea56161ce8dd0f1b08dd20f5c43c05551d8f71dd748621774741a1627e5

Observation e5adc283-597b-4b00-82c4-e31932f82b47 · outbound

This paper cites AI safety via debate.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback AI safety via debate

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.663769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.663769Z digest=sha256:382c81c52657cd872afec616b15852e6d49d120e63c78358ff5332e3ae41d866

Observation 211cf569-56ac-4ca5-b6fb-5609d1d9d85e · outbound

This paper cites Language Models (Mostly) Know What They Know.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Language Models (Mostly) Know What They Know

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.669363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.669363Z digest=sha256:e469ac951b5fbe878f83456f3cde998c9e18f28fa5af46d23a6d7f5685aedd4c

Observation 19245a6b-857e-4a22-b0fa-cae434bd318c · outbound

This paper cites Algorithms for inverse reinforcement learn- ing.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Algorithms for inverse reinforcement learn- ing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:17.071285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-06T15:14:16.684460Z digest=sha256:21338c237d597a949ccda93954b9d22b81c67da5bab7624b2169ca0f4c1bdcbc

Observation 436d7b01-70a3-4706-abef-90a5af555369 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.693748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.693748Z digest=sha256:52345d9df71e185684f91f4144cbeeaf4e10d485b86b8e8c9674e970d08c0d48

Observation 06b5a3b5-ecfd-448d-91dd-f902350515fc · outbound

This paper cites A new system-wide diversity measure for recommendations with efficient algorithms.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback A new system-wide diversity measure for recommendations with efficient algorithms

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:14:16.828662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-06T15:14:16.703719Z digest=sha256:9514a52754a0bfbfe17a2f35fbff2c10ee124593dcf7d1bf3870476867c6bc19

Observation 0fe0c28a-d246-4698-861e-37cba15abb66 · outbound

This paper cites The Fates of Merging Supermassive Black Holes and a Proposal for a New Class of X-Ray Sources.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback The Fates of Merging Supermassive Black Holes and a Proposal for a New Class of X-Ray Sources

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:14:16.803484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-06T15:14:16.708646Z digest=sha256:cc71e07ae2bea2ebd9f85bb22b7846266b1db517fe56209a4c84b33f81ef4683

Observation 02197dd9-d428-456a-b78c-87ad2378f729 · outbound

This paper cites Red Button.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Red Button

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:17.034115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-06T15:14:16.727994Z digest=sha256:6813fd6db8ad7fd2fbad1caf6c436d57338e46c78db77830dc162438448a5990

Observation 8a1a899e-d56f-4157-9daa-0fef7608adf1 · outbound

This paper cites Discovering Latent Knowledge in Language Models Without Supervision.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Discovering Latent Knowledge in Language Models Without Supervision

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.689102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.689102Z digest=sha256:22120d25440a42162886c9fd7035611761e979e24ef7c1e0bc1fba6155f76db6

Observation dd39c946-5aac-4e79-aa74-8da336e6e031 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.648067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.648067Z digest=sha256:c075b500dd45ca3848ebb178e58186faa9baa5501d3acc9e39613d1c05f4b42d

Observation ad369e9c-2031-4c53-8021-f4544014d216 · outbound

This paper cites Supervising strong learners by amplifying weak experts.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Supervising strong learners by amplifying weak experts

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.652974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.652974Z digest=sha256:c752dc274c2d8ac595e042665491f18348a7672f57010d053f31ff5721d36f82

Observation a1494bd8-4baf-4426-8e33-f0d8c6450a55 · outbound

This paper cites Improving alignment of dialogue agents via targeted human judgements.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Improving alignment of dialogue agents via targeted human judgements

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.657949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.657949Z digest=sha256:0f7b421aaaa7daa9fcd5c6fc83498c9a3bd7fcff9943a0d75db111f83c0ccb4f

Observation 4cbba20d-fc0e-45f2-833b-801d7b7fbe1b · outbound

This paper cites Self-critiquing models for assisting human evaluators.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Self-critiquing models for assisting human evaluators

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.698792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.698792Z digest=sha256:04f05451d1935dcb8acc28a9a7eff75cd694f340347d314f3378ad3149856ab7

Observation 858244ef-3017-4256-819a-15683edcdaf7 · outbound

This paper cites Spin-Spin Coupling at Small $x$: Worm-Gear and Pretzelosity TMDs.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Spin-Spin Coupling at Small $x$: Worm-Gear and Pretzelosity TMDs

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.713331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.713331Z digest=sha256:d4f7c7ede808365ad942bfa79020a4226d7d138de9452b3d3bd09e04bb76d27f

Observation d9c4c7f2-95ab-46a5-925c-e1e001df020f · outbound

This paper cites Teaching language models to support answers with verified quotes.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Teaching language models to support answers with verified quotes

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.679333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.679333Z digest=sha256:5c4df11d9fa6a1e3adc17ee857b5393fa1761d8f576da1fc41b48141fe7684a8

Observation 0719c205-ff88-44cc-a248-71620dc3f487 · outbound

This paper cites Pebble: Feedback- efficient interactive reinforcement learning via relabeling experience and un- supervised pre-training.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Pebble: Feedback- efficient interactive reinforcement learning via relabeling experience and un- supervised pre-training

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:17.087493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-06T15:14:16.674730Z digest=sha256:97fb86282cee0fa9b2e6416920e58a3ca69b4c1d678e152d5f382fc0c385235a

Observation 8029b33b-822d-4a61-b8c2-f69458cf70f9 · outbound

This paper cites Red button,.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Red button,

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:17.052660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-06T15:14:16.720775Z digest=sha256:4ca4763b38794cc7b0f9f2f4dd19a6d9ee849760cd7fc2ca5e158e3fe1822382

Pith citing papers

No inbound Pith citation observations are available.