Pith. sign in

Paper Citation Record · LEDGER

Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2312.04127.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.04127 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T00:12:04.906792Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T02:20:44.610149Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1959d544-707d-43b3-b06f-152cdf5255e3 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.612725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:84ed8964f64a6237f9eec47607ef5552e301553ffbb83075cf316c1983b6b2c6

Observation e11f8152-d47f-4d6c-b415-1b67239037d0 · inbound

Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models cites this paper.

Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T00:12:04.906792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:12:04.906792Z digest=sha256:1782e2bc6971033c97608afc46997ed8e7dc110d14beccaf3bd19ef6b5c52283

Observation 3e918652-9044-4560-b738-db78cebd1452 · inbound

Adversarial Preference Learning for Robust LLM Alignment cites this paper.

Adversarial Preference Learning for Robust LLM Alignment Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:21.021039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:21.021039Z digest=sha256:2d236198dc1d566c911187c2d573b4517dc3dfe55c413f18ded786a160210158

Observation f2de84f1-870e-469e-b6f7-00b81abe62f6 · inbound

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models cites this paper.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.014753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.014753Z digest=sha256:e5257f83ecb846926f5abc5eba7d4c05b58b1bd7ea6fd7eecad05825143748cc

Observation e935f0d9-eb2b-400a-96d5-da4ed0bfeee9 · inbound

Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models cites this paper.

Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:25:55.064809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:25:55.064809Z digest=sha256:cbee22f083631a3c8cf7ff1e9346571f1053cad47e2ad23f2688a4dad7e227fd

Observation 2861fd44-e50f-4d4a-8c03-1ae8259f8c1b · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:33.736025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:33.736025Z digest=sha256:0c103937b49c9338d1520d444c67911d7368aa11d64e989d7dab89151eb58e7f

Observation fe38cfcf-af02-4ad5-ae37-7a429babd349 · inbound

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint cites this paper.

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T23:10:21.282394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:10:21.282394Z digest=sha256:bb30c60765493c32056debbf4a59b754db1ac8046facee170976acd72a34aa0f

Observation af156b49-2e8e-48a1-b665-2b82b8b86dd3 · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.456061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.456061Z digest=sha256:560e59b7fe7953b16bf5845ab2c8730ca2e2d9a83af9e640a814b0a6ac6baac7

Observation f8f1cedc-1993-45e7-b0cd-7c3cfd65c564 · inbound

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems cites this paper.

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:16.893553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:16.893553Z digest=sha256:d8187517d69821202c28f20c7421b496318d914bb9972a89b99a71ea730d668c