Pith. sign in

Paper Citation Record · LEDGER

Should LLM Safety Be More Than Refusing Harmful Instructions?

As of 9 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2506.02442.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02442 v2

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:27:47.482807Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact3
  • verified fuzzy2
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4f626177-35eb-4ab9-bd5b-a0c7274a307b · outbound

This paper cites GPT-4 Technical Report.

Should LLM Safety Be More Than Refusing Harmful Instructions? GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.122877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.122877Z digest=sha256:cc47e4ca93d84102bd27e92ecda0a5aaea1ba4f0c29af26c19c8559fcc608a25

Observation 52404c4d-5d6a-471e-a625-cda4dadbb19f · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:50.178348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.129536Z digest=sha256:365595d7061faea27177ab9f5f32c64e85ba31114179db1a6683f894f5ce2d7f

Observation ae55402e-b8c8-4b91-8a4b-b978302c4afb · outbound

This paper cites Detecting Language Model Attacks with Perplexity.

Should LLM Safety Be More Than Refusing Harmful Instructions? Detecting Language Model Attacks with Perplexity

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.137122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.137122Z digest=sha256:f4b7ed1eb09778c5e1392ee8801775d936bfb55cabeea88b75e6d7141e56e191

Observation 04624ff3-35b9-404b-ab65-c8b96582c804 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? Gemini: A Family of Highly Capable Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.144854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.144854Z digest=sha256:b208ded027331e8d4ae1cd4030b72a98b8c1f974b2b753aaf76a5eb4da3c3447

Observation cc9f277e-c466-46c5-ac76-a412e39331ef · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:50.095009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.152801Z digest=sha256:82bc939c93224ff86f2f89b9cdbbf3b85512f1468ea79c55e27a687428358a1a

Observation a2c1901f-66e2-4c83-a48e-6db1d4fb6c5c · outbound

This paper cites Jailbreaking Large Language Models with Symbolic Mathematics.

Should LLM Safety Be More Than Refusing Harmful Instructions? Jailbreaking Large Language Models with Symbolic Mathematics

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.159009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.159009Z digest=sha256:158234b256326e6e17b4182075f48d4b01d7fbf0d861692fb5fb0dddc89f91eb

Observation f11eae64-55f7-45ac-b29d-bf93198671f2 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.166556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.166556Z digest=sha256:0f412f08ba5b60801990c97c5d6dc130d42a5914782afc561531bc2558588d84

Observation bbb14b9f-4929-4411-a53d-120cf8f8b928 · outbound

This paper cites Pappas, Florian Tram\` e r, Hamed Hassani, and Eric Wong.

Should LLM Safety Be More Than Refusing Harmful Instructions? Pappas, Florian Tram\` e r, Hamed Hassani, and Eric Wong

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:27:49.928300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.175424Z digest=sha256:dea52ee89d6582a7cf15c4618824e2ec83de47ca48d1a5041dba38cc9b28779e

Observation c43bbd36-f338-4770-bffb-980105f6b432 · outbound

This paper cites Pappas, Florian Tram \`e r, Hamed Hassani, and Eric Wong.

Should LLM Safety Be More Than Refusing Harmful Instructions? Pappas, Florian Tram \`e r, Hamed Hassani, and Eric Wong

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:27:49.777409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.182198Z digest=sha256:d6e40c254853b52c3714239b8297520295e1aacb4413cad7527415a97fc96296

Observation 3ecfed4b-90af-4563-bbef-b163fb4546b3 · outbound

This paper cites Recent Advances in Attack and Defense Approaches of Large Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? Recent Advances in Attack and Defense Approaches of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.188977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.188977Z digest=sha256:df2369d6415d78646f8fded043a0b01de8f5148858af038f00b1184c6dc040f0

Observation 5ca62124-816a-4f1f-9737-e4b7ec72e693 · outbound

This paper cites Attacks, Defenses and Evaluations for LLM Conversation Safety: A Survey.

Should LLM Safety Be More Than Refusing Harmful Instructions? Attacks, Defenses and Evaluations for LLM Conversation Safety: A Survey

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.195714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.195714Z digest=sha256:502d1c599aa800b3e00ab5cc0785e8844a01161ef23d3091ae81311b2ac0ed02

Observation c6b7078e-91a8-422d-af32-06c9c369cce8 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 12

Resolution
verified exact
doi, observed 2026-08-07T11:27:47.649065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.204744Z digest=sha256:cc3198af63af212ca641513feca446082b54c65ef8e76aab89b9820fc0e18900

Observation 72ae24e7-ab71-481b-9ca2-3158ebd550e2 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.647650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.213100Z digest=sha256:b1558be575653aecb63297159faa6d5126332b13e4dcc2e88f70607c40b0aa9c

Observation 88b1e8b6-d140-46b4-a892-d0a8b718f7a3 · outbound

This paper cites Unsupervised Cipher Cracking Using Discrete GANs.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unsupervised Cipher Cracking Using Discrete GANs

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:27:48.698390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.227011Z digest=sha256:8b2445c6967653ce35f1056f53c22c39d59e23a56a73b33ca5e46f374d5c99ad

Observation 30ab4d27-9b68-4ebd-98b4-569097c9ddc8 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Should LLM Safety Be More Than Refusing Harmful Instructions? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.235852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.235852Z digest=sha256:5e93a049892bc906d32ee6cc0f389823581259000839bb93e86b6614af6d9c08

Observation 567d72f0-94fd-42bf-9c2d-ba662c14ee19 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.501436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.242254Z digest=sha256:5d4fb7894a85ed52093e4f178b638b43276ed0fc28a181a0782ed4fd1a096948

Observation b07cb090-4083-40a9-ae2b-5094165514ae · outbound

This paper cites competency.

Should LLM Safety Be More Than Refusing Harmful Instructions? competency

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.250751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.250751Z digest=sha256:45b05a9cc58a292928c00a95e823f41e39667c189010e008897d127a04987358

Observation f21550f7-2f32-4419-af9e-5c182045b666 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.259907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.259907Z digest=sha256:52db62f0f8faf890786394abd4585d7b2e5580e4364456f7a67c8db795143aab

Observation 5444761c-5f45-403b-93b7-032d835e6c19 · outbound

This paper cites Endless Jailbreaks with Bijection Learning.

Should LLM Safety Be More Than Refusing Harmful Instructions? Endless Jailbreaks with Bijection Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.272740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.272740Z digest=sha256:51baebf52ed2733032de3fa1c49267bc86b5082708855ea709414fd68fd361d0

Observation 88b28cec-7931-4bd0-ab13-fd3bf806ad74 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Should LLM Safety Be More Than Refusing Harmful Instructions? Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.280600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.280600Z digest=sha256:022a8f486d657a0567fc755c4d6f0c899bfc000bb0cb988ef252a52e6200f82f

Observation d449b9ff-6312-494c-9fce-64e26117404f · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.286452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.286452Z digest=sha256:7966d4d7d9461304f381e96af0906f902a3d12c231949d15a1d8da6dfb4cd3cf

Observation 73344d01-9c0e-4c79-b7c9-733d12031035 · outbound

This paper cites Mistral 7B.

Should LLM Safety Be More Than Refusing Harmful Instructions? Mistral 7B

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.295083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.295083Z digest=sha256:7168e01345e86a00cdc5f23fcc6a94d27e2cfd87788e0052d02b2dc3369ccd1f

Observation 7093f919-ff68-48f4-8443-7131b57c8455 · outbound

This paper cites ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs.

Should LLM Safety Be More Than Refusing Harmful Instructions? ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.302318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.302318Z digest=sha256:95d992061697bac128fe998b75478829bd38e6002d50f10b4d013137215ab270

Observation d5aaf8e2-e118-45dc-a18d-0d865ef72af0 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.310189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.310189Z digest=sha256:86066c5039e74bf1e020603326e39e126ae501d58158ed1788a168e18955755e

Observation 25be7c48-77af-4c52-8fb1-e932c0bf2cd7 · outbound

This paper cites CipherBank: Exploring the Boundary of LLM Reasoning Capabilities through Cryptography Challenges.

Should LLM Safety Be More Than Refusing Harmful Instructions? CipherBank: Exploring the Boundary of LLM Reasoning Capabilities through Cryptography Challenges

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.317124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.317124Z digest=sha256:559ed692137d8e852b03aa198bb00ae165d9e38c789f91ba6f74733939043415

Observation d3919f68-8a83-4a68-bfdd-33df43cb7dea · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.325604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.325604Z digest=sha256:fc970df95425cde9826f4aac3bb5223013f3faed49c73d0e846963ed61715bbc

Observation b091d066-6201-4946-a08c-2e603dcaceda · outbound

This paper cites CodeChameleon: Personalized Encryption Framework for Jailbreaking Large Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? CodeChameleon: Personalized Encryption Framework for Jailbreaking Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.338065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.338065Z digest=sha256:467cbc671d193fc1bcad16dc4355c3cd63a4f5076418a63732c69249baa9b77f

Observation 642d42cb-e567-4d53-81b7-f1669b2ff556 · outbound

This paper cites Benchmarking Large Language Models for Cryptanalysis and Side-Channel Vulnerabilities.

Should LLM Safety Be More Than Refusing Harmful Instructions? Benchmarking Large Language Models for Cryptanalysis and Side-Channel Vulnerabilities

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.346115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.346115Z digest=sha256:2891aee3e30c537d90ab7c95d43cdabb6d118373fc1841f1d9b7f36403982bfd

Observation 855c9b16-6bc5-494d-a596-913432d3e6df · outbound

This paper cites SaRO: Enhancing LLM Safety through Reasoning-based Alignment.

Should LLM Safety Be More Than Refusing Harmful Instructions? SaRO: Enhancing LLM Safety through Reasoning-based Alignment

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.354047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.354047Z digest=sha256:306c2473d553f35f41f01bbd26264577c278064b871071022fed18c52ff8bf00

Observation e83d5f35-8bbc-474d-88e1-43945ded8fd0 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 30

Resolution
verified exact
raw_fallback, observed 2026-08-07T11:27:48.173546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.369437Z digest=sha256:fe0512a5269170e7539410db8be7f6d76022baed84de0dbdc5df16b2c061a00b

Observation e64ce843-07d1-48f9-a367-3428c3f2529b · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.375342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.375342Z digest=sha256:578ad77189a715858c67cc9e989cc709467583f3407dd56733c50e84b0df8863

Observation ccd256c9-e1ec-46b8-ada0-20f9e9743acd · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.381480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.381480Z digest=sha256:49932c2f86f62990b2c25972567a149a6e64328289037a9a4edb43f008b6b6de

Observation cf9f0039-c899-41fd-96e7-deba8fd39018 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.399243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.386841Z digest=sha256:e50d5d9790973d578bd583f251d9996787414d9e58127af0dcb1037691857c4f

Observation 97a19b08-9b36-452a-b4f3-21651ea0bb56 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.393410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.393410Z digest=sha256:607cd5ab2774e35a1aaa94fb6ec6d4fabe907bcf5cfdafae63f5c38fc0f033e7

Observation 9b9643af-eeed-44d3-97cf-291528fbb9bd · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.272487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.399564Z digest=sha256:500a4dff2694a96bd1fce6f5d4208a947362bb790955f60c2b00c3afd432c0a8

Observation 28416321-807b-4036-978f-a1716094620e · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? LLaMA: Open and Efficient Foundation Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.405477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.405477Z digest=sha256:5ffa5078d01242b2d86ef2c4c735ea0d6e9b8b58cff1d89d0d816776e695f0b9

Observation 3299d20c-9e50-4068-9f5b-e6512fbb7019 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.412052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.412052Z digest=sha256:beed9ad21f6ef53ee393637b02eb771f5db2215e3144927eea29896cbeb4fc17

Observation 83349e79-5cd6-4017-b2b7-2d15e981583f · outbound

This paper cites Dai, and Quoc V Le.

Should LLM Safety Be More Than Refusing Harmful Instructions? Dai, and Quoc V Le

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.418202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.418202Z digest=sha256:47dcf973c05523e7c8460ea379ced28fdab17b825d86825fab9f14d8351fbaf8

Observation 210a72bd-66f7-41e1-b106-a2140787b95f · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.424467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.424467Z digest=sha256:42ee276695e95031c2e96079ba6f5f28ae589344a73e736131bb78a5cc33922f

Observation cfd2488b-dc86-4db2-b031-d07166648d10 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.018405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.430319Z digest=sha256:a36e1e120922f746471601a173b91afae1681f85f4bbed63f4b381ac2ccaf3d9

Observation 3f819ca3-cc5c-413f-bf67-8f459f50fc71 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.435792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.435792Z digest=sha256:d2595c8d96adab20b702fa0db9e6a7d4389ae61c8312d8da1c0c4b6f24c6f5aa

Observation 309629fa-7449-4a7c-9598-8fb71354e27a · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.443024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.443024Z digest=sha256:1cbafbde4d7fe7ce700f94a5464120762132ae9af041f0792d3dd15508a2ff89

Observation b54cedb5-02a6-4160-b10b-f12f9acbbc3f · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.449114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.449114Z digest=sha256:ca4a3358649f766720a67517c6d83fda279ec3135737b1edc3e985cc73c33e6b

Observation c32352f6-5eb5-42a8-8ed4-8f8deedc95e8 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.454913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.454913Z digest=sha256:c753d1789855965e7b897395d8cdd6040b17729eb9d53b06692ebbcbf1320394

Observation 4e57b8c6-f9cd-4a4b-b2f0-07c74fd00440 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.461431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.461431Z digest=sha256:8057a7e17549e44522b7f03abb0b574c609ee879aa34f6489b3489601bc554ed

Observation 86e2506d-b829-49ed-8cf4-cd50cc697115 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Should LLM Safety Be More Than Refusing Harmful Instructions? BERTScore: Evaluating Text Generation with BERT

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.467267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.467267Z digest=sha256:12617f815af47395ff656355f4e38329a2c2fcf8e8ed91447880ccfa4665fd7f

Observation 411437fe-d19a-48b6-a933-82d84bff3c92 · outbound

This paper cites online" 'onlinestring :=.

Should LLM Safety Be More Than Refusing Harmful Instructions? online" 'onlinestring :=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.475035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.475035Z digest=sha256:569915f204fb04945a052366abd264b95a1dd323062db767aabc7a6b234dc2a9

Observation 4d16e431-eaa6-413b-a14e-88b605b2ba9d · outbound

This paper cites write newline.

Should LLM Safety Be More Than Refusing Harmful Instructions? write newline

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.482807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.482807Z digest=sha256:45332fec8f382be53997e09fc48ec1c8e56eed1ed5f4a2e37940343013329c13

Pith citing papers

No inbound Pith citation observations are available.