Pith. sign in

Paper Citation Record · LEDGER

The Superalignment of Superhuman Intelligence with Large Language Models

As of 13 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2412.11145.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11145 v2

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:18:17.886907Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5ee5ceee-70c8-4a54-9594-ff6f9165c7e9 · outbound

This paper cites Large Language Model Alignment: A Survey.

The Superalignment of Superhuman Intelligence with Large Language Models Large Language Model Alignment: A Survey

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.764452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.764452Z digest=sha256:8287956bde8216911adba1c9ac6ef40ae55e8f9be56b0248e68883d070a2d188

Observation 6a4a4bff-479e-4ce9-a2de-474314a45475 · outbound

This paper cites Statistical Rejection Sampling Improves Preference Optimization.

The Superalignment of Superhuman Intelligence with Large Language Models Statistical Rejection Sampling Improves Preference Optimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.773756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.773756Z digest=sha256:2094153c3ed08fb69900835458d4e5de5f957543e4dcc302056bed068d6828f9

Observation ab1c205b-8504-432f-98d1-6245523cdc74 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

The Superalignment of Superhuman Intelligence with Large Language Models KTO: Model Alignment as Prospect Theoretic Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.778579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.778579Z digest=sha256:9b15fd68dbfdf5ec3a32fd119b745855825fdcecb3111c6df1f784f7dada38bd

Observation d32a8254-b368-4624-95fa-9520e269f644 · outbound

This paper cites Measuring Progress on Scalable Oversight for Large Language Models.

The Superalignment of Superhuman Intelligence with Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.783327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.783327Z digest=sha256:c848ff93d9f88671db2300e87feb9ac3ac2d7be73a208392b629b4c82963ba29

Observation 16b50d46-0267-4805-a866-16ef8e1779be · outbound

This paper cites Supervising strong learners by amplifying weak experts.

The Superalignment of Superhuman Intelligence with Large Language Models Supervising strong learners by amplifying weak experts

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.788272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.788272Z digest=sha256:ca29490fc5537ac605ca02b9fa3a39fdd79dca25a3688618d382067e43b0b182

Observation a28e9c60-0409-4a0e-ade6-96397527604a · outbound

This paper cites Scalable agent alignment via reward modeling: a research direction.

The Superalignment of Superhuman Intelligence with Large Language Models Scalable agent alignment via reward modeling: a research direction

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.792623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.792623Z digest=sha256:c7ab22446599c91c24d32bfd5bf8285b72570eeee5e4389b91b1cf07aaf504db

Observation 4f0ea86e-646b-42c1-8dd0-c87bb1a3d15c · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

The Superalignment of Superhuman Intelligence with Large Language Models Constitutional AI: Harmlessness from AI Feedback

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.796886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.796886Z digest=sha256:1868d12f7252c32829df801f1acd07c7297e3b46202e07aad1689d7f8ebc8d6a

Observation 3dd39d37-f285-4817-9a59-e0666b752e76 · outbound

This paper cites AI safety via debate.

The Superalignment of Superhuman Intelligence with Large Language Models AI safety via debate

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.801169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.801169Z digest=sha256:f569f0219b0c1733669b8cdbc02a04fc45a6de7d0730416c8710fd99a45b88e3

Observation 59ef8334-a65e-445d-b030-20965e2b20c8 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

The Superalignment of Superhuman Intelligence with Large Language Models Training Verifiers to Solve Math Word Problems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.810199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.810199Z digest=sha256:924b6817fa85c84844c3cefe0df053d7e57db3feeb32afede62735fa316aba06

Observation b66629ff-e189-4614-bdf9-3f3a95ab5386 · outbound

This paper cites LLM Critics Help Catch LLM Bugs.

The Superalignment of Superhuman Intelligence with Large Language Models LLM Critics Help Catch LLM Bugs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.820149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.820149Z digest=sha256:bf6d60010825f438a07671c359e9532dc3e85cf386e7ecfd4475c8d826313636

Observation 2c67b252-b1f5-4a07-bc7a-bbc985a2c63a · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

The Superalignment of Superhuman Intelligence with Large Language Models Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.824191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.824191Z digest=sha256:cb8b8cefa40eacaa472012f38874e7aca841dd47ac2c4ff5ed31ac6c89257ac4

Observation 2708c8af-192b-4a68-93db-24b239e5a1e5 · outbound

This paper cites AutoDetect: Towards a unified framework for automated weakness detection in large language models.

The Superalignment of Superhuman Intelligence with Large Language Models AutoDetect: Towards a unified framework for automated weakness detection in large language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.293578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T15:18:17.829537Z digest=sha256:bb294f6d9418af5d683e56125fded5db9c1e856c037e16d91c0f0e36c83b40e0

Observation ecd67083-f9ee-4f72-baab-447e193754a4 · outbound

This paper cites Language Models Learn to Mislead Humans via RLHF.

The Superalignment of Superhuman Intelligence with Large Language Models Language Models Learn to Mislead Humans via RLHF

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.834330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.834330Z digest=sha256:ce4c9d6a7e811b6f01cfa66be0789f4c5a23a69ead17db895f2e5e39745b2a58

Observation 1047eca7-37bc-493c-b977-30c9fb9dcd0b · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

The Superalignment of Superhuman Intelligence with Large Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.838574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.838574Z digest=sha256:7979dd6300ac786e2076f744f25063be4b17668d1649c6d8be9b72edc310542e

Observation 2db461c9-03b7-44bb-9755-b95547d1b5d9 · outbound

This paper cites Unveiling the implicit toxicity in large language models.

The Superalignment of Superhuman Intelligence with Large Language Models Unveiling the implicit toxicity in large language models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.279193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T15:18:17.842695Z digest=sha256:3d2296a953defe744986f2175f757fb3838dd010c00c42c4a26e5e080ba80bfe

Observation 2d76fa22-507b-429e-a715-bcfd64e27215 · outbound

This paper cites Improving Reward Models with Synthetic Critiques.

The Superalignment of Superhuman Intelligence with Large Language Models Improving Reward Models with Synthetic Critiques

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.846870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.846870Z digest=sha256:8a74daf8c3ed9237854ddfabbed342801f7eeb00e676ae7770c2d8189b60db22

Observation d9196234-2fa9-4c5c-9a18-9db170704218 · outbound

This paper cites Reinforcement Learning for Generative AI: A Survey.

The Superalignment of Superhuman Intelligence with Large Language Models Reinforcement Learning for Generative AI: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.851686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.851686Z digest=sha256:8e7d1bffc08283267f71ec737bf2ce3e35b7fd105cda21e88f93b4d0ce4508fc

Observation c6b3b56d-02ac-4c2f-a0a9-7bef81be63a4 · outbound

This paper cites Learning to refine with fine-grained natural language feedback.

The Superalignment of Superhuman Intelligence with Large Language Models Learning to refine with fine-grained natural language feedback

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.265397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T15:18:17.856738Z digest=sha256:df4a8d14238a4263461cc749b7eb3973d0005e33046d02fbb642d1159c62fba5

Observation 81eac04b-19eb-454d-b5f6-73eb1e5e90f8 · outbound

This paper cites Improving Model Factuality with Fine-grained Critique-based Evaluator.

The Superalignment of Superhuman Intelligence with Large Language Models Improving Model Factuality with Fine-grained Critique-based Evaluator

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.860973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.860973Z digest=sha256:f275d8066efb88ab9a4c33393daf4aef07d761a524098a86d8f950efb9ce7a8f

Observation 5db97a81-6283-4170-99ba-26eda00bbc34 · outbound

This paper cites Bayesian calibration of win rate estimation with LLM evaluators.

The Superalignment of Superhuman Intelligence with Large Language Models Bayesian calibration of win rate estimation with LLM evaluators

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.251679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T15:18:17.864978Z digest=sha256:624780d504299eaab866468d4a856afc62eec7b4e71c1506697e2df2735c9f2e

Observation 717fbb25-ced1-4ed4-b551-15a149fe3a90 · outbound

This paper cites Trust or Escalate: LLM Judges with Provable Guarantees for Human Agreement.

The Superalignment of Superhuman Intelligence with Large Language Models Trust or Escalate: LLM Judges with Provable Guarantees for Human Agreement

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.868903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.868903Z digest=sha256:906a6b45136c341088ea4c76da8f378c8824fe1d68a85e7ea2100cbebc54d81b

Observation dfbecd98-f46b-4f7b-92fb-2dc7d38a88b7 · outbound

This paper cites JudgeLM: Fine-tuned Large Language Models are Scalable Judges.

The Superalignment of Superhuman Intelligence with Large Language Models JudgeLM: Fine-tuned Large Language Models are Scalable Judges

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.873601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.873601Z digest=sha256:b35b51ecafe40779eb41303dd1d6d73560216e0308c145f5ed9a589532309353

Observation f631249e-2027-428f-a169-0f54032ef4b2 · outbound

This paper cites Self-critiquing models for assisting human evaluators.

The Superalignment of Superhuman Intelligence with Large Language Models Self-critiquing models for assisting human evaluators

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.877823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.877823Z digest=sha256:e65c5de3399f16b720e9488d662dc8191c1818b90d5e3a0ad802c91c0c7a62d7

Observation d28700c0-7ecd-4477-8c5e-6f7da6145206 · outbound

This paper cites Lm vs lm: Detecting factual errors via cross examination.

The Superalignment of Superhuman Intelligence with Large Language Models Lm vs lm: Detecting factual errors via cross examination

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.237875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T15:18:17.882052Z digest=sha256:10e94659ee0d8572952b554006a8db2a2b00bbca81bb391e33e239e024dc0a3b

Observation 9b7f6cc6-638b-4e6f-8427-d88ade9038f5 · outbound

This paper cites Panacea: Pareto Alignment via Preference Adaptation for LLMs.

The Superalignment of Superhuman Intelligence with Large Language Models Panacea: Pareto Alignment via Preference Adaptation for LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.886907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.886907Z digest=sha256:43ac56cc1347b185f7c7747ff5d67cbcd3f16c9c872cfcb40f3de48ddac60456

Observation 9c2a3476-5879-499e-9109-0c8b9ff6b1a9 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

The Superalignment of Superhuman Intelligence with Large Language Models Evaluating Large Language Models Trained on Code

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.815334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.815334Z digest=sha256:32e3a7865a44296f48b26e307556b70b37fc288b75e6ed5cc468418652f1cac9

Observation 5c7ea3fc-7624-491e-8f72-99b0ad5f281e · outbound

This paper cites A Survey of Large Language Models.

The Superalignment of Superhuman Intelligence with Large Language Models A Survey of Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.753976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.753976Z digest=sha256:1d8099f5bf1ff349300b049557b5193381067cfc8abeea4ecdce288727fda714

Observation b6ced059-cdeb-46d0-bcca-3f7ffa1b1e4b · outbound

This paper cites Proximal Policy Optimization Algorithms.

The Superalignment of Superhuman Intelligence with Large Language Models Proximal Policy Optimization Algorithms

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.768904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.768904Z digest=sha256:96bb87dcfc16110114f3f0f6a259a889e039e7f7eda806941faaf44a683184b4

Observation 9719a814-22c5-47cf-b1ad-ab859268a7d5 · outbound

This paper cites Model evaluation for extreme risks.

The Superalignment of Superhuman Intelligence with Large Language Models Model evaluation for extreme risks

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.759688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.759688Z digest=sha256:752bc4c76682ee67e72ce1dc36f1d1fdbe9153ebb16c52752fcd2bb0aaf1d624

Pith citing papers

No inbound Pith citation observations are available.