Pith. sign in

Paper Citation Record · LEDGER

Fundamental Limitations of Alignment in Large Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2304.11082.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.11082 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:58:34.926982Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

44
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bc298fbf-da94-4833-8e0d-3f5871613c44 · inbound

Jailbroken: How Does LLM Safety Training Fail? cites this paper.

Jailbroken: How Does LLM Safety Training Fail? Fundamental Limitations of Alignment in Large Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-14T18:17:42.828873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T18:17:42.752997Z digest=sha256:ca5fbd87165ede4488510456dfa91e3af5f4b64538d0933a01423d47d8cac8e6

Observation 3efae1f4-54d7-4687-aafa-7b082e8c0467 · inbound

Universal and Transferable Adversarial Attacks on Aligned Language Models cites this paper.

Universal and Transferable Adversarial Attacks on Aligned Language Models Fundamental Limitations of Alignment in Large Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-24T07:44:08.470910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T07:42:09.112946Z digest=sha256:c1e0e95d98d3b60358411d41ef3d6ed3e241d6a38524101cae0a4dcdd9632601

Observation f6879a9a-010c-4b07-8abb-41580cb58f24 · inbound

Salamandra Technical Report cites this paper.

Salamandra Technical Report Fundamental Limitations of Alignment in Large Language Models

Reference 215

Resolution
unresolved
no resolver link, observed 2026-08-08T04:58:34.926982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:58:34.926982Z digest=sha256:29f51daddf39ceaf170a86967c0720359969f8eecc57bf9cf04f648868fcee5c

Observation 5e43b040-e749-4864-b1e7-84ea0d259d51 · inbound

Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race cites this paper.

Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race Fundamental Limitations of Alignment in Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:19.496083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:13:19.496083Z digest=sha256:9a57b39e1d5d2f6d8a667f86bcde274048793aeb1eacef4d8dac00ec13024f54

Observation 1885a77b-acf8-4ed0-bd2d-6a19a4d708c2 · inbound

JavelinGuard: Low-Cost Transformer Architectures for LLM Security cites this paper.

JavelinGuard: Low-Cost Transformer Architectures for LLM Security Fundamental Limitations of Alignment in Large Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:25.526186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:41:25.526186Z digest=sha256:11ff71492a40c3685c86e38622e679f72de227e6b6c2421e76f8e5edaedb40af

Observation 6decc58e-6b91-4b12-af93-4bdc9bc1e436 · inbound

Hyperbolic Deep Learning for Foundation Models: A Survey cites this paper.

Hyperbolic Deep Learning for Foundation Models: A Survey Fundamental Limitations of Alignment in Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T14:54:36.557710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:54:36.557710Z digest=sha256:a391038734548b3b51196b56b672a25e8fc3c4171a10b23e31adb0ccf41d0201

Observation 560af3c0-046c-40eb-90de-d44ddc6a6834 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Fundamental Limitations of Alignment in Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.796947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.796947Z digest=sha256:a86b81b578afd8920eb0a8053c04628c2425e4b9ae5dc29a0407c8a35535938a

Observation 5e0dd17a-3fe9-4cf7-97a3-67fc29cff162 · inbound

Robust AI Security and Alignment: A Sisyphean Endeavor? cites this paper.

Robust AI Security and Alignment: A Sisyphean Endeavor? Fundamental Limitations of Alignment in Large Language Models

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T22:58:38.228561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T22:58:29.306435Z digest=sha256:759948a00cea9188e909c0c627ec4389d26d5524455784065237203d30d900b6

Observation 6a98c7df-c782-4da9-8ac3-581311a95014 · inbound

Plausible Patients, Impossible Populations: Auditing Epidemiological Fidelity in Large Language Model Mental Health Simulations cites this paper.

Plausible Patients, Impossible Populations: Auditing Epidemiological Fidelity in Large Language Model Mental Health Simulations Fundamental Limitations of Alignment in Large Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:06:18.723757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T06:04:24.077276Z digest=sha256:6b495735b05dbbec74f48ee29100c4d2af0410d88e83c1b6be64adc9dd87ad60

Observation 5d7d489b-4ec9-4eb0-8d06-8dab12957bd0 · inbound

Latent Personality Alignment: Improving Harmlessness Without Mentioning Harms cites this paper.

Latent Personality Alignment: Improving Harmlessness Without Mentioning Harms Fundamental Limitations of Alignment in Large Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:51:43.581200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T01:46:49.586630Z digest=sha256:941f29c8547bfab07bda3e48a23724532ebabf381a2c4f7a9fe832570d2bcc8f

Observation a9fe88b3-3c82-4db0-9281-54d53f7f3910 · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Fundamental Limitations of Alignment in Large Language Models

Reference 206

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:17:55.556966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:9fd3ae3932debd8734dd6e3100cd3a4ad01e0f44713dd87c42a45d89ef670be4

Observation 9b99e65a-6ad6-4fdd-bcf4-65cbeffdebe5 · inbound

Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders cites this paper.

Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders Fundamental Limitations of Alignment in Large Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:49:10.130331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T22:48:54.238767Z digest=sha256:730704c514faa763e513dd534ac362a5d66f4e9110678068e75398d3507f46e3

Observation 3de5c74b-9fdd-42c4-8631-8a55023605de · inbound

Trusted Weights, Treacherous Optimizations? Optimization-Triggered Backdoor Attacks on LLMs cites this paper.

Trusted Weights, Treacherous Optimizations? Optimization-Triggered Backdoor Attacks on LLMs Fundamental Limitations of Alignment in Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:49:35.636729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:45:35.079192Z digest=sha256:31a91b48da24e061fa2bced8381c4cd32ce261b666ac1709bdbf30506d9f3d2f

Observation e70bd3cf-98ea-48d2-84aa-478bb6a88973 · inbound

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible cites this paper.

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible Fundamental Limitations of Alignment in Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:34:01.928075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T22:28:02.493124Z digest=sha256:12223aafaf56e622fbaca3912d239687e0e140d433ae616cc05772a2e02f5850

Observation 669a96b2-01be-4161-880d-11824b753b97 · inbound

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible cites this paper.

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible Fundamental Limitations of Alignment in Large Language Models

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-02T13:17:57.695515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:17:57.695515Z digest=sha256:e4f322c852110af7ddd7bbb1c87fca0f971fdb5efe97d5b944dfb80e56169ccb

Observation 7fab3428-33eb-4396-bbc3-6788746fba32 · inbound

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms cites this paper.

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms Fundamental Limitations of Alignment in Large Language Models

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:32:53.150041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T00:26:54.019256Z digest=sha256:101c5c3d6b530a41686d7dc6711d962260d414242fb24a28733d1a82221eb27d

Observation 8fc14387-4f84-447c-8574-00d7ebe85e8d · inbound

Confused ChatGPT: Cross-App Context Poisoning via First-Party APIs cites this paper.

Confused ChatGPT: Cross-App Context Poisoning via First-Party APIs Fundamental Limitations of Alignment in Large Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:32:35.659788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T18:55:45.465260Z digest=sha256:2a710fee27de1c1235e78db6e59cfd9f17c27ed3bbca10b1a068693c7999ba4d

Observation def98b02-acfb-4a38-bbab-17645d140d20 · inbound

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy cites this paper.

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy Fundamental Limitations of Alignment in Large Language Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:57:25.953535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T18:36:44.265273Z digest=sha256:2d1f6be5252bf9afd499ab3ca8127db7997dd8539ae627d620f875abccb2161a

Observation 92bd060d-0b28-4cbd-9348-c6f9896922f7 · inbound

On The Effectiveness-Fluency Trade-Off In LLM Conditioning: A Systematic Study cites this paper.

On The Effectiveness-Fluency Trade-Off In LLM Conditioning: A Systematic Study Fundamental Limitations of Alignment in Large Language Models

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T11:18:03.179730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T09:40:48.736006Z digest=sha256:fd88e2f944f80ceba8ea462cd425dda8054c3374248b7439045bd9618820bac1

Observation d459a3d0-42eb-458d-81b9-10a3b9baff94 · inbound

Test-Time Scaling via Error Localization cites this paper.

Test-Time Scaling via Error Localization Fundamental Limitations of Alignment in Large Language Models

Reference 132

Resolution
unresolved
no resolver link, observed 2026-08-01T07:28:31.376035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:28:31.376035Z digest=sha256:1637b8a79507cd443a1e376459ef3f31c7fdf7e572220a230d9c0a19948fa2d4