Pith. sign in

Paper Citation Record · LEDGER

Towards medical AI misalignment: a preliminary study

As of 9 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2505.18212.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18212 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:52:17.039027Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9352bc8f-88eb-4d94-9ab0-317079dd4a2e · outbound

This paper cites Mistral Large system card, 2024.

Towards medical AI misalignment: a preliminary study Mistral Large system card, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:19.086894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:15.110635Z digest=sha256:476f1f2bb02d2a8513dca7a08775d97dc1e9f4a72290bb00de707db01f01cb17

Observation ff977c06-a41e-472f-8c02-836e6d22019e · outbound

This paper cites Many-shot jailbreaking.

Towards medical AI misalignment: a preliminary study Many-shot jailbreaking

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.896857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:15.159722Z digest=sha256:0f7d32c8c9ae0fe01aefeefd096b6654e5b1461423ae1995d99a7f1339587818

Observation 7977d1d4-7841-4643-87d2-a3097260b1df · outbound

This paper cites LLMs Will Always Hallucinate, and We Need to Live With This.

Towards medical AI misalignment: a preliminary study LLMs Will Always Hallucinate, and We Need to Live With This

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.208059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.208059Z digest=sha256:62fd640350c5d6c8603d48668f4c4ed37bcbac90a93ed33721f23a6e62d8dd11

Observation 06ef82a5-dede-480c-814b-4f8b94c662b2 · outbound

This paper cites Superhuman performance of a large language model on the reasoning tasks of a physician.

Towards medical AI misalignment: a preliminary study Superhuman performance of a large language model on the reasoning tasks of a physician

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.279718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.279718Z digest=sha256:9ca8f3c4425a80565bf8906835c89d542294d75a92dc16a814b05ef079f1f4bd

Observation b34e61b5-4d31-4f35-b218-b322820944d3 · outbound

This paper cites an unresolved cited work.

Towards medical AI misalignment: a preliminary study Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:52:18.793916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:15.339144Z digest=sha256:85133e7da95e500018d3fb755c9d257bcfd8344ec849a3e470ca4ccfa87962ab

Observation 7327ce62-76ff-4af4-9824-b7f52304ddd7 · outbound

This paper cites Providing explanations via the EQR argument scheme.

Towards medical AI misalignment: a preliminary study Providing explanations via the EQR argument scheme

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.681839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:15.416513Z digest=sha256:4d7da6eadc7686a429b4738bb612eff30a3884fb54c440099b53bd28b8085553

Observation 9220165c-0f83-4a54-8e6b-180be994e0c6 · outbound

This paper cites JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs.

Towards medical AI misalignment: a preliminary study JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.472625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.472625Z digest=sha256:2a618c89271f23de876a03d130ea6c4d81790595f1adabc09e05317d33d66114

Observation 37eb964d-83b0-48d2-b96b-bb82039b6d59 · outbound

This paper cites MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots.

Towards medical AI misalignment: a preliminary study MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.531093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.531093Z digest=sha256:a4b01af8adf54b7685a7a5c9fedf552790d6ffd5217ce4e257cc54933433cb5b

Observation 3049beb4-c141-45ea-9275-ea79311b5366 · outbound

This paper cites Bias and fairness in large language models: A survey.

Towards medical AI misalignment: a preliminary study Bias and fairness in large language models: A survey

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.573013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:15.605146Z digest=sha256:6e6260a4d6b31cb63230a3d0f656fe219ec02cadf62c622ffa85e48ace9d1fc7

Observation ff7d0451-dc62-4404-a5d4-9158b7c11cdc · outbound

This paper cites Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks.

Towards medical AI misalignment: a preliminary study Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:52:17.372837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:15.678133Z digest=sha256:bfed1cbb24be86b6bff2e6c4eed148039d65db64d37dbaaf01fee5adc265ee75

Observation 6971aa2b-d808-4905-80cf-ec45d8af6108 · outbound

This paper cites Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation.

Towards medical AI misalignment: a preliminary study Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.741594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.741594Z digest=sha256:3c2a0bb099ccd4c05880d95e86bcd7ec34f8cafca4d4c6a070f10839e2a12ff5

Observation 70aeb414-4a52-44f0-b4ba-29477f2bd247 · outbound

This paper cites Quack: Automatic jailbreaking large language models via role-playing.

Towards medical AI misalignment: a preliminary study Quack: Automatic jailbreaking large language models via role-playing

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.445224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:15.798209Z digest=sha256:dc17618f73f9d660cad69634996a9664543697b484c770ecf913764317248a4e

Observation 90d5d5d0-731c-4275-a359-c3044dc2bfcf · outbound

This paper cites Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models.

Towards medical AI misalignment: a preliminary study Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.873668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.873668Z digest=sha256:0402d32ef5dda4c881166c0eab1588d768e582391129fa0084809655a077a2a6

Observation 8e8cf2df-b10e-489a-ac96-bc15025f4f03 · outbound

This paper cites Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character.

Towards medical AI misalignment: a preliminary study Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.961141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.961141Z digest=sha256:e6ddcb9952250a234302e7c56a03d2499c2acd3972f993c37f31021e977e3423

Observation 3e851a72-e28b-4d1c-bb81-0ff04d53827e · outbound

This paper cites Capabilities of GPT-4 on Medical Challenge Problems.

Towards medical AI misalignment: a preliminary study Capabilities of GPT-4 on Medical Challenge Problems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.041569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.041569Z digest=sha256:2ca0c355bf233f2114d1cd46891aef6541495d0cb675a6451ce3449d2f0bb63f

Observation 4d07d32f-4715-4ecc-8781-1ee494b9f0d7 · outbound

This paper cites Gpt-4 system card, 2023.

Towards medical AI misalignment: a preliminary study Gpt-4 system card, 2023

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.296299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:16.110736Z digest=sha256:6ae9bddf7f556295ca654c7f4167a158469d6d44df83c54952b8e89f433ad7fd

Observation 9ba7b60a-2bc1-40a7-bfc0-5773a6fcc5f7 · outbound

This paper cites Openai o1 system card, 2024.

Towards medical AI misalignment: a preliminary study Openai o1 system card, 2024

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.191940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.191940Z digest=sha256:087d224293bd339f0cea1c47e0cf833f568ff88ffd6ebe5a4718c5d52e737e3e

Observation 8cf758c1-5ddb-4a39-ae5c-9072ee497a65 · outbound

This paper cites A course in game theory.

Towards medical AI misalignment: a preliminary study A course in game theory

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.257112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.257112Z digest=sha256:d5232cc7e40639ee3843e4a22dca8285d58b9f8115d2924c2db5211c53c2611f

Observation 25bc1c2d-9085-4e86-9300-7d3ec4eb2c7c · outbound

This paper cites A survey of machine learning in healthcare.

Towards medical AI misalignment: a preliminary study A survey of machine learning in healthcare

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.160725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:16.315130Z digest=sha256:247eeec45b1313163f8d0a83bc633b949f6368809ea4b2a1239fdc7fb4183761

Observation e44bd1bf-5e81-4a86-9e1e-6becb2d75a0e · outbound

This paper cites Do anything now.

Towards medical AI misalignment: a preliminary study Do anything now

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.041166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:16.367791Z digest=sha256:41c5ba21f6702f724c518a88eaa5e5d1f7e455f0a2c7c13e648bdc5331d01d10

Observation dfbda8b3-0a04-46b4-877a-8dc8a485c303 · outbound

This paper cites Large Language Models Encode Clinical Knowledge.

Towards medical AI misalignment: a preliminary study Large Language Models Encode Clinical Knowledge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.426651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.426651Z digest=sha256:a6041027b52b00452b60eb5040b65531b22ce8ba5168179665d12e7edc470440

Observation 13a9fa32-6156-45a9-845b-a22c9d5ae15a · outbound

This paper cites A survey of clinicians’ views of the utility of large language models.

Towards medical AI misalignment: a preliminary study A survey of clinicians’ views of the utility of large language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:17.891732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:16.478581Z digest=sha256:8ac7c45bf4577ddd0c07b363880ac5b222dc1fd631d2c7e0b143c5173e515be4

Observation 8713db44-0cbd-4303-8aa2-6a47b9f376c6 · outbound

This paper cites Introducing Gemini 2.0: our new AI model for the agentic era, 2024.

Towards medical AI misalignment: a preliminary study Introducing Gemini 2.0: our new AI model for the agentic era, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:17.774025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:16.552335Z digest=sha256:3fe24601c6427607537e2bc0676cd37ff98e61b158973220eb344e23ae703437

Observation f6347c9a-ad51-4a26-a54e-653de97859c9 · outbound

This paper cites DeepSeek-V3 Technical Report.

Towards medical AI misalignment: a preliminary study DeepSeek-V3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.605845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.605845Z digest=sha256:fe9cc095f5093e26f5145acef2c6ed2cebf797c7a1c0c0b6b2e62d17ae9e9ffc

Observation 286874bc-a130-402c-bb90-fa2dd56076c1 · outbound

This paper cites Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems, 36, 2024.

Towards medical AI misalignment: a preliminary study Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems, 36, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.679218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.679218Z digest=sha256:0d7881c6c3519673d42e7c7e279bdc4e7b4baf0fab4b3d53cc7c6a923f5c2892

Observation 20a2cb91-98eb-4da2-b0d2-041042033da8 · outbound

This paper cites Evaluating the use of large language models to provide clinical recommendations in the emergency department.

Towards medical AI misalignment: a preliminary study Evaluating the use of large language models to provide clinical recommendations in the emergency department

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:17.624248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:16.731189Z digest=sha256:8c88be4f302363b68a67247936ad17d86f5279c9ed6d4e4dafdc8f18481a299d

Observation d4530a51-d4bf-4b9a-83d9-c2641f95cfed · outbound

This paper cites A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models.

Towards medical AI misalignment: a preliminary study A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.793630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.793630Z digest=sha256:722be39679dc89612fff5d9d80dff98e7a865102d14c22f858b3884eddf2ad08

Observation 096e1651-6c3e-4a49-852d-d06bf3dd0d7d · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

Towards medical AI misalignment: a preliminary study Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.858220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.858220Z digest=sha256:5d5ee400457bd37f5a10e9a42ddf56af49b9bfc6398692f50bac3c85c5958122

Observation 93bf1d20-5cec-4421-a34e-59cce3628d5e · outbound

This paper cites Low-Resource Languages Jailbreak GPT-4.

Towards medical AI misalignment: a preliminary study Low-Resource Languages Jailbreak GPT-4

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.928485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.928485Z digest=sha256:4df3c90fcc272c1a98378729cbbb8db84bd53f27703e5b9148156410445c108e

Observation 4ef5a3be-d2eb-4649-8aae-6c959e60940c · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Towards medical AI misalignment: a preliminary study Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:17.523453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:52:16.986295Z digest=sha256:4db3d9a831ff3c72fdadc956becb8c2e1661c87e40cd0b4ecc0e9f253339e9d1

Observation fc8677b1-1b50-4912-83bc-a503187c2d4a · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Towards medical AI misalignment: a preliminary study Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:17.039027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:17.039027Z digest=sha256:ef54b77316a2f92ca440132c2f6beeb2d184d4cc4fe43f80bcc0b58ac562a0d2

Pith citing papers

No inbound Pith citation observations are available.