Pith. sign in

Paper Citation Record · LEDGER

Fake Alignment: Are LLMs Really Aligned Well?

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2311.05915.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.05915 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:54:33.285486Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T07:37:04.295513Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ceb64e9b-0b37-4ff0-902b-abb35181231c · inbound

GPAI Evaluations Standards Taskforce: Towards Effective AI Governance cites this paper.

GPAI Evaluations Standards Taskforce: Towards Effective AI Governance Fake Alignment: Are LLMs Really Aligned Well?

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-12T15:54:33.285486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:54:33.285486Z digest=sha256:3b9d19306873f1f6315ea8521bd5d0992d51a728ba1563189e4dd22795a64460

Observation e31f9482-b5a0-44ba-ab4e-7991955b234c · inbound

Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models cites this paper.

Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models Fake Alignment: Are LLMs Really Aligned Well?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:52.256199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:52.256199Z digest=sha256:4f4796063dc157542a41f320f69d574f64962a7f7f3bd12394fb330f9ddf5b45

Observation 623bc94d-6277-4646-93e3-e520db0653df · inbound

Bridging Distribution Shift and AI Safety: Conceptual and Methodological Synergies cites this paper.

Bridging Distribution Shift and AI Safety: Conceptual and Methodological Synergies Fake Alignment: Are LLMs Really Aligned Well?

Reference 192

Resolution
unresolved
no resolver link, observed 2026-08-07T13:05:35.411291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:05:35.411291Z digest=sha256:b704d6dcd13beeba98c15f73e1196e630523b67e99a987f77b5032c469d05fa1

Observation a8efab55-a054-434d-8ef5-7b8c02e87b1a · inbound

Why do AI agents communicate in human language? cites this paper.

Why do AI agents communicate in human language? Fake Alignment: Are LLMs Really Aligned Well?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:49.757309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:49.757309Z digest=sha256:ac1d84dcef3636d04de97dc9e4ca05d6cfecec09903e22a1a17faf7bad91537e

Observation 8ec9bd09-a478-4f06-8507-756bbe4b9125 · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges Fake Alignment: Are LLMs Really Aligned Well?

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.023624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.023624Z digest=sha256:d1603491fd7c4ab824c75bc72e6d46dfcebb535d97d5b67e050d7f7b2df2f1e0

Observation f3cd0a44-9e8a-4d43-a9a4-9279b82d2539 · inbound

When Choices Become Risks: Safety Failures of Large Language Models under Multiple-Choice Constraints cites this paper.

When Choices Become Risks: Safety Failures of Large Language Models under Multiple-Choice Constraints Fake Alignment: Are LLMs Really Aligned Well?

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:37:04.297016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T07:36:10.174861Z digest=sha256:cc5c78828313de85aaa3c8391d3339f3a8e5e5cbd65b709038ba61f89d12f5fb