Pith. sign in

Paper Citation Record · LEDGER

Towards medical AI misalignment: a preliminary study

As of 19 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2505.18212.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18212 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:52:17.039027Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9352bc8f-88eb-4d94-9ab0-317079dd4a2e · outbound

This paper cites Mistral Large system card, 2024.

Towards medical AI misalignment: a preliminary study Mistral Large system card, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:19.086894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:15.110635Z digest=sha256:b4df05fc39c73624f3deaf53938229626e8fc4267c93b50b61ee2086366255ee

Observation ff977c06-a41e-472f-8c02-836e6d22019e · outbound

This paper cites Many-shot jailbreaking.

Towards medical AI misalignment: a preliminary study Many-shot jailbreaking

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.896857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:15.159722Z digest=sha256:d0f48d82effb5eaba515f35e23b1a6d7ec3e0ae2902ee9e6fd3bc6dd049e9be6

Observation 7977d1d4-7841-4643-87d2-a3097260b1df · outbound

This paper cites LLMs Will Always Hallucinate, and We Need to Live With This.

Towards medical AI misalignment: a preliminary study LLMs Will Always Hallucinate, and We Need to Live With This

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.208059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.208059Z digest=sha256:5c694d52fb46d5335a613b4f8bcdde9ac6a4178bbfc6e9cf9c28519a6172e6a3

Observation 06ef82a5-dede-480c-814b-4f8b94c662b2 · outbound

This paper cites Superhuman performance of a large language model on the reasoning tasks of a physician.

Towards medical AI misalignment: a preliminary study Superhuman performance of a large language model on the reasoning tasks of a physician

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.279718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.279718Z digest=sha256:3641dc764ce85c8a42291845d8d34dfa4a386529e5499b9e338a5812d5b7f253

Observation b34e61b5-4d31-4f35-b218-b322820944d3 · outbound

This paper cites an unresolved cited work.

Towards medical AI misalignment: a preliminary study Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:52:18.793916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:15.339144Z digest=sha256:135f7da56e26c617b045a73c6d3b7d64877e437fe09ae9589e0891b9cd9f21d6

Observation 7327ce62-76ff-4af4-9824-b7f52304ddd7 · outbound

This paper cites Providing explanations via the EQR argument scheme.

Towards medical AI misalignment: a preliminary study Providing explanations via the EQR argument scheme

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.681839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:15.416513Z digest=sha256:a35377c5e9d85f131aea56b185fa47d64e9d5ac31fc724a2936d2d8d6637c399

Observation 9220165c-0f83-4a54-8e6b-180be994e0c6 · outbound

This paper cites JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs.

Towards medical AI misalignment: a preliminary study JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.472625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.472625Z digest=sha256:a7e5e25588ef19932de7366d6b8529af560a63d676644d3ac8cc43e9b1ed3c46

Observation 37eb964d-83b0-48d2-b96b-bb82039b6d59 · outbound

This paper cites MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots.

Towards medical AI misalignment: a preliminary study MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.531093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.531093Z digest=sha256:8f1893fa1fe5476b9d09877e2add0213b7331ad27e06b1a85151e4166c75e72c

Observation 3049beb4-c141-45ea-9275-ea79311b5366 · outbound

This paper cites Bias and fairness in large language models: A survey.

Towards medical AI misalignment: a preliminary study Bias and fairness in large language models: A survey

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.573013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:15.605146Z digest=sha256:7198cd5dc34c19616773137836312e85842c902a67c7db3203792a3b5ffcc05f

Observation ff7d0451-dc62-4404-a5d4-9158b7c11cdc · outbound

This paper cites Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks.

Towards medical AI misalignment: a preliminary study Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:52:17.372837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:15.678133Z digest=sha256:eee11c960e10e2d8a3801a71c8bef5f30a1564092f5798a7a8122fa5fb42debc

Observation 6971aa2b-d808-4905-80cf-ec45d8af6108 · outbound

This paper cites Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation.

Towards medical AI misalignment: a preliminary study Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.741594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.741594Z digest=sha256:8673c3c667246b8a70598276e1e7e6d2222e9ae86186d8803b0ce19c3fdfbd2d

Observation 70aeb414-4a52-44f0-b4ba-29477f2bd247 · outbound

This paper cites Quack: Automatic jailbreaking large language models via role-playing.

Towards medical AI misalignment: a preliminary study Quack: Automatic jailbreaking large language models via role-playing

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.445224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:15.798209Z digest=sha256:83641196988f1174fc1657a70dc1e0d6f229918acd873a845b2438ad60541e40

Observation 90d5d5d0-731c-4275-a359-c3044dc2bfcf · outbound

This paper cites Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models.

Towards medical AI misalignment: a preliminary study Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.873668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.873668Z digest=sha256:29605999ee360bf661e9c4d809c45edbf54c2f316d8918a4b2155dc5810cf774

Observation 8e8cf2df-b10e-489a-ac96-bc15025f4f03 · outbound

This paper cites Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character.

Towards medical AI misalignment: a preliminary study Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.961141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.961141Z digest=sha256:2c836ba8d2a2e9947163650c538cca8470c14735e50c5783b9f4bee21c071334

Observation 3e851a72-e28b-4d1c-bb81-0ff04d53827e · outbound

This paper cites Capabilities of GPT-4 on Medical Challenge Problems.

Towards medical AI misalignment: a preliminary study Capabilities of GPT-4 on Medical Challenge Problems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.041569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.041569Z digest=sha256:23c22be6a54d8c74295182f752cfcc7f15b7bd34eca792e1f59973994eba11e7

Observation 4d07d32f-4715-4ecc-8781-1ee494b9f0d7 · outbound

This paper cites Gpt-4 system card, 2023.

Towards medical AI misalignment: a preliminary study Gpt-4 system card, 2023

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.296299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:16.110736Z digest=sha256:86a413670b61027f64293f5651a1a13f4080f0e756c1718a4087b5a433fecb81

Observation 9ba7b60a-2bc1-40a7-bfc0-5773a6fcc5f7 · outbound

This paper cites Openai o1 system card, 2024.

Towards medical AI misalignment: a preliminary study Openai o1 system card, 2024

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.191940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.191940Z digest=sha256:17286c3fa0f2b41188e573eb98cc50de61356516edcfc9569b90d771f7edf529

Observation 8cf758c1-5ddb-4a39-ae5c-9072ee497a65 · outbound

This paper cites A course in game theory.

Towards medical AI misalignment: a preliminary study A course in game theory

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.257112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.257112Z digest=sha256:0d7f1cad5471824efcff711dde05172ea0ca50e6b58f077fa038ddda2215a09b

Observation 25bc1c2d-9085-4e86-9300-7d3ec4eb2c7c · outbound

This paper cites A survey of machine learning in healthcare.

Towards medical AI misalignment: a preliminary study A survey of machine learning in healthcare

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.160725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:16.315130Z digest=sha256:6dc7ad01d7390170cf9054039d0e5035f1b51a59c2fa518c0c89630bcb30cd94

Observation e44bd1bf-5e81-4a86-9e1e-6becb2d75a0e · outbound

This paper cites Do anything now.

Towards medical AI misalignment: a preliminary study Do anything now

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:18.041166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:16.367791Z digest=sha256:41e555ebb3755e9f0086c1848005bda224aa73839159d85ef6fa8ad92316fcd2

Observation dfbda8b3-0a04-46b4-877a-8dc8a485c303 · outbound

This paper cites Large Language Models Encode Clinical Knowledge.

Towards medical AI misalignment: a preliminary study Large Language Models Encode Clinical Knowledge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.426651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.426651Z digest=sha256:f770e8d41f638787338b53f984bce8bef659600926ca7b9d2cd9ec3abf23dca0

Observation 13a9fa32-6156-45a9-845b-a22c9d5ae15a · outbound

This paper cites A survey of clinicians’ views of the utility of large language models.

Towards medical AI misalignment: a preliminary study A survey of clinicians’ views of the utility of large language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:17.891732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:16.478581Z digest=sha256:08702028c080bfcea4ae2e772cf3487f8b98441ce55f196638311d6b1c830762

Observation 8713db44-0cbd-4303-8aa2-6a47b9f376c6 · outbound

This paper cites Introducing Gemini 2.0: our new AI model for the agentic era, 2024.

Towards medical AI misalignment: a preliminary study Introducing Gemini 2.0: our new AI model for the agentic era, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:17.774025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:16.552335Z digest=sha256:6e80c97f3b70249785d20c298f35cd802c26673b3e58bf07ec1939db4a43201d

Observation f6347c9a-ad51-4a26-a54e-653de97859c9 · outbound

This paper cites DeepSeek-V3 Technical Report.

Towards medical AI misalignment: a preliminary study DeepSeek-V3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.605845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.605845Z digest=sha256:e03521d107c00d75ecf59fe3d70599193544a9af7a0514e909f8eb47060c8f3c

Observation 286874bc-a130-402c-bb90-fa2dd56076c1 · outbound

This paper cites Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems, 36, 2024.

Towards medical AI misalignment: a preliminary study Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems, 36, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.679218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.679218Z digest=sha256:74f16b793de032d3dfe9414af372535a7e1eeadc97d87a1a15c6d53e2c5bf876

Observation 20a2cb91-98eb-4da2-b0d2-041042033da8 · outbound

This paper cites Evaluating the use of large language models to provide clinical recommendations in the emergency department.

Towards medical AI misalignment: a preliminary study Evaluating the use of large language models to provide clinical recommendations in the emergency department

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:17.624248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:16.731189Z digest=sha256:535c259700d5be591b7bcf94d0ec7554bd646b4d1f1f2e96ecaf7484966f03fe

Observation d4530a51-d4bf-4b9a-83d9-c2641f95cfed · outbound

This paper cites A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models.

Towards medical AI misalignment: a preliminary study A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.793630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.793630Z digest=sha256:7b073911e750c4f985016676adf7e78aedc9be2792e246255a0875a415e894ad

Observation 096e1651-6c3e-4a49-852d-d06bf3dd0d7d · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

Towards medical AI misalignment: a preliminary study Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.858220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.858220Z digest=sha256:dc954d11de8c952cc04605381983807e04a6f36591af749f90c7cb7042a62890

Observation 93bf1d20-5cec-4421-a34e-59cce3628d5e · outbound

This paper cites Low-Resource Languages Jailbreak GPT-4.

Towards medical AI misalignment: a preliminary study Low-Resource Languages Jailbreak GPT-4

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.928485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.928485Z digest=sha256:b35d21f9360ba0bba3c68745635818912c71ec9e95fb014c7b2c763683fa2498

Observation 4ef5a3be-d2eb-4649-8aae-6c959e60940c · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Towards medical AI misalignment: a preliminary study Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:52:17.523453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:52:16.986295Z digest=sha256:0ceff98d8daf1ec9937f34cf321999a49b83dfada8e2404b77128ae55722d6de

Observation fc8677b1-1b50-4912-83bc-a503187c2d4a · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Towards medical AI misalignment: a preliminary study Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:17.039027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:17.039027Z digest=sha256:1de195a0a918c5124a57eba20e6c17a8acd406a7b6e27b8dacaa1ff657b4bf2c

Pith citing papers

No inbound Pith citation observations are available.