Pith. sign in

Paper Citation Record · LEDGER

Extracting and Understanding the Superficial Knowledge in Alignment

As of 16 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 1 inbound Pith citation observation for arXiv:2502.04602.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04602 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T22:14:55.954664Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:52:18.183674Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:52:20.114728Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f588d168-5325-409a-81e8-e998300cb5b6 · outbound

This paper cites an unresolved cited work.

Extracting and Understanding the Superficial Knowledge in Alignment Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:14:56.149662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T22:14:55.926806Z digest=sha256:efcdd2b424d30930450a0bba8c5bb7960ed8bb16d6909462f5df4512a7f0ca3c

Observation dabecbaf-3516-4d76-9400-7a3dd2a88527 · outbound

This paper cites Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!.

Extracting and Understanding the Superficial Knowledge in Alignment Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T22:14:55.901982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:14:55.901982Z digest=sha256:bfcaac0f3f3e78898aee58eaade20be61f0ccaaac1c69f8b8886eae36092bdcb

Observation 7e22fdeb-01a3-428d-8908-1974101a9539 · outbound

This paper cites an unresolved cited work.

Extracting and Understanding the Superficial Knowledge in Alignment Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:14:56.119552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T22:14:55.935645Z digest=sha256:0fe1dc48fa024e7cd0f48e2061332fd748926cc48433c3dfbb1f14d0ca021f84

Observation ffcff278-451e-4bac-8f59-623b8d7df4e0 · outbound

This paper cites Crowdsourcing Multiple Choice Science Questions.

Extracting and Understanding the Superficial Knowledge in Alignment Crowdsourcing Multiple Choice Science Questions

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T22:14:55.911615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:14:55.911615Z digest=sha256:1274ab69425ba729c6fcf03570d64733d7c4ebf2740a41b3cc2a6fe92ee1a010

Observation 4b204672-b5be-4f54-8fa9-98deaeae6810 · outbound

This paper cites an unresolved cited work.

Extracting and Understanding the Superficial Knowledge in Alignment Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:14:56.163664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T22:14:55.922032Z digest=sha256:b760aad705c695117a6d196ffec4b3e90abae1b7b00f62b8943d6bbcdf2a3740

Observation 6bd857ca-973d-4508-9395-4236ba25ad48 · outbound

This paper cites an unresolved cited work.

Extracting and Understanding the Superficial Knowledge in Alignment Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:14:56.135303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T22:14:55.931132Z digest=sha256:2d54ab3e7b27416e98e21f5f37dc7add54dd081b4dc7bf93a5f1a4fc8cb3c481

Observation 15447bff-6be4-428c-b145-ab70b759eda1 · outbound

This paper cites So, the alarm rang a total of 4 + 12 + 6 = 22 times today.

Extracting and Understanding the Superficial Knowledge in Alignment So, the alarm rang a total of 4 + 12 + 6 = 22 times today

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:14:56.105425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T22:14:55.940083Z digest=sha256:49d21fd3c00a3480c5367b2ec9500e89b74c3b75aa7b482f987a37b3d2eb30fd

Observation e109e982-23d9-4b79-88cc-ad94712f609a · outbound

This paper cites an unresolved cited work.

Extracting and Understanding the Superficial Knowledge in Alignment Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:14:56.090229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T22:14:55.944473Z digest=sha256:7701829287738b69f3da3de64c5ed8237f6f0062b21aec5a745006c2978b6ca1

Observation e17a306a-1089-4e0c-8a20-b2772fc3a0cb · outbound

This paper cites an unresolved cited work.

Extracting and Understanding the Superficial Knowledge in Alignment Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:14:56.075714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T22:14:55.949669Z digest=sha256:2eede477adf2ee5a5c3c4bd92b58af603ea8326eb1b7f963b453d0d456a3d6f1

Observation 339af629-73cd-4bde-811d-1c4be1170c04 · outbound

This paper cites So, in total, the Alarm rang for 4 + 12 + 6 = 22 seconds.

Extracting and Understanding the Superficial Knowledge in Alignment So, in total, the Alarm rang for 4 + 12 + 6 = 22 seconds

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:14:56.061185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T22:14:55.954664Z digest=sha256:11eccb0b49bd9754e9307616b625ae0979e0b79930ad7ac08c70862e43b2c479

Observation 4f9fe4d0-2dfd-4328-bc38-f1508cbc812f · outbound

This paper cites Why Should Adversarial Perturbations be Imperceptible? Rethink the Research Paradigm in Adversarial NLP.

Extracting and Understanding the Superficial Knowledge in Alignment Why Should Adversarial Perturbations be Imperceptible? Rethink the Research Paradigm in Adversarial NLP

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T22:14:55.896568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:14:55.896568Z digest=sha256:dd939f32d7dbe590d084661ddddb465e3be08774695dc3ad4125da26af8dcb8b

Observation da3d686f-eead-4a74-9faf-6f57156e6ac3 · outbound

This paper cites Reinforcement Learning from Diverse Human Preferences.

Extracting and Understanding the Superficial Knowledge in Alignment Reinforcement Learning from Diverse Human Preferences

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T22:14:55.916487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:14:55.916487Z digest=sha256:d183fee071e84f0916eacf1d7dbe902f3df816f357f00f5e9dc893476636855a

Observation 034065b0-b058-4ffa-bd56-af61e0436141 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Extracting and Understanding the Superficial Knowledge in Alignment Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T22:14:55.906875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:14:55.906875Z digest=sha256:0b697208e83407063d14e97ae63ec0b75c3c01998335c685cdc8b76fe21d757c

Pith citing papers

Observation 6165a161-d4dd-4bd9-8a99-aead95d924fd · inbound

What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding cites this paper.

What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding Extracting and Understanding the Superficial Knowledge in Alignment

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:52:20.121950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T05:52:18.183674Z digest=sha256:7c34a084ff892263de12938ffbe29ff0316f89b062364a8b59307d62e75f4e0a