Pith. sign in

Paper Citation Record · LEDGER

Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2312.04127.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.04127 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:21.021039Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T02:20:44.610149Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1959d544-707d-43b3-b06f-152cdf5255e3 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.612725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:6f584c5dae3e85577ae7806dc8d9850315100574177a3fb7c205258df670678b

Observation 3e918652-9044-4560-b738-db78cebd1452 · inbound

Adversarial Preference Learning for Robust LLM Alignment cites this paper.

Adversarial Preference Learning for Robust LLM Alignment Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:21.021039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:21.021039Z digest=sha256:1373eb9c4893c4924a39aee231f00af294a695ab55a6059c75db97ddc224588e

Observation f2de84f1-870e-469e-b6f7-00b81abe62f6 · inbound

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models cites this paper.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.014753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.014753Z digest=sha256:c5ec5bd14eabd56ae839a79b45cf35d0325de1e9f56e603357785c0f2985f04e

Observation e935f0d9-eb2b-400a-96d5-da4ed0bfeee9 · inbound

Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models cites this paper.

Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:25:55.064809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:25:55.064809Z digest=sha256:72b971a62a9f05986252df6a738a39f678bc5703b489f5f676132611efb8ed3b

Observation 2861fd44-e50f-4d4a-8c03-1ae8259f8c1b · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:33.736025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:33.736025Z digest=sha256:dd1fbc979455b3f79a27ccccd638526764507f40021cbf7d5e3b2539c0a7f379

Observation fe38cfcf-af02-4ad5-ae37-7a429babd349 · inbound

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint cites this paper.

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T23:10:21.282394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:10:21.282394Z digest=sha256:4d4f9c4322b769ab2b7bdf1050f53bdb04ef1cbd13e059a3df67f760e7d6dc7d

Observation af156b49-2e8e-48a1-b665-2b82b8b86dd3 · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.456061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.456061Z digest=sha256:44fde91b994054e3802a736b1d7c6e87a840176df3789667555d65d964642003

Observation f8f1cedc-1993-45e7-b0cd-7c3cfd65c564 · inbound

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems cites this paper.

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:16.893553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:16.893553Z digest=sha256:421438a0b6663bcaa5aa1ff82b546df5aba4b6d894c3a9acbbe8a296d1a60dda