Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T07:01:49.222325Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2607.03968.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T07:01:49.222325Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3bb8f63f-d30f-43be-8d61-dfd7a46afe36 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Using large language models for multi-level commit message generation for large diffs,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27799e65-5fc3-4abd-88aa-be0363a837f1 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Summarize me: The future of issue thread interpretation,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9518fea1-e73d-4eb0-a150-017d8673981a · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Large language models for software engi- neering: A systematic literature review,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 297d5a61-f2b6-44d8-b3e6-ff7019c41760 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f296076a-ad16-4c32-99e5-948ede4b3c51 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents A systematic evaluation of large language models of code,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 234c1bed-9d20-4312-9117-a73c2a63ec83 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ffd0578-f33b-4a04-9f97-d11162ec0621 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Language models for code completion: A practical evaluation,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88a1b7ef-679e-479b-aa47-94ad31f5bff8 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents A review on code generation with llms: Application and evaluation,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4b6b2a4-983a-46fb-a945-a3a224793c59 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 580bb6d7-6b73-421a-b14f-02db6794d56d · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c56be8a3-ac77-4531-843a-6fb3f7279177 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 971b9bd0-b9c8-46dc-b1f3-4551475613a2 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Don’t say no: Jailbreaking llm by suppressing refusal,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e5b7cd7-4993-4e93-bcf2-529fa95ed843 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Endless jailbreaks with bijection learning,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3693395a-6f3e-407b-8a2b-e1e7be03c29b · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Many-shot jailbreaking,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d4fb4e5-36bb-4b48-bfb7-34d3ce874c58 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Codeattack: Revealing safety generalization challenges of large lan- guage models via code completion,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7ad9aa9-b190-4f8b-9b69-648dfb57e760 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Great, now write an article about that: The crescendo{Multi-Turn}{LLM}jailbreak attack,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a4e9452-c72a-4b12-a65a-3abdf8e56def · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Code red! on the harmfulness of applying off-the-shelf large language models to programming tasks,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1098af1d-4af5-4470-8d30-6e25479fdf3c · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents What Breaks When LLMs Code? Characterizing Operational Safety Failures of Agentic Code Assistants
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6676761b-75c8-41a0-a11d-bbf4dc26c2fc · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d39c16c6-ac1c-47f6-94b1-5c34437337a1 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak and guard aligned language models with only few in-context demonstrations,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 934d8f74-9ba6-467b-bac7-946e5e2e6054 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Smoke and Mirrors: Jailbreaking LLM-based Code Generation via Implicit Malicious Prompts
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d2c6a6f-88f7-4a31-a2cf-dda07a7df6d6 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Redcode: Risky code execution and generation benchmark for code agents,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba10b06a-18b6-4678-b764-e04350b519dd · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e38be30-be24-4324-9734-978f2d789f73 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Refusal-Trained LLMs Are Easily Jailbroken As Browser Agents
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 777998ec-968d-4417-bf33-fbafa7dcbc5b · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Agentdojo: A dynamic environment to evaluate prompt injection attacks and defenses for llm agents,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d6729c1-235c-4217-bd9f-6b08cfff4630 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Os-harm: A benchmark for measuring safety of computer use agents, 2025,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d576a44-fcf0-4f74-82a8-e2b1bf542b52 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Redcoder: Automated multi-turn red teaming for code llms,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d631095a-1211-4997-afa3-eabd4b9fdfd8 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Breaking the Code: Security Assessment of AI Code Agents Through Systematic Jailbreaking Attacks
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 646b76ec-4243-4da9-901e-2c53bc1219ef · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Sema: Simple yet effective learning for multi-turn jailbreak attacks,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b78426c-7ced-42f2-b964-4f713e454373 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6bdaf64-4287-4eb9-9133-2f2413721191 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Llms as evaluators: A novel approach to commit message quality assessment,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfe038c8-ed6b-4e43-a486-271dfd3ca83a · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28a407e4-338f-44fd-bf39-6c4af7adee6b · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents A coin flip for safety: Llm judges fail to reliably measure adversarial robustness,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92b849b1-c7f8-4bfe-ba88-60e431f2db42 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents JAILJUDGE: A Comprehensive Jailbreak Judge Benchmark with Multi-Agent Enhanced Explanation Evaluation Framework
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a200fc6-1537-442d-873b-706e30c97462 · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Llms cannot reliably judge (yet?): A comprehensive assessment on the robustness of llm-as-a-judge,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b62673e9-7d84-45f7-a38b-f6d5e0294d6a · outbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Large Language Models are Inconsistent and Biased Evaluators
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.