Pith. sign in

Paper Citation Record · LEDGER

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

As of 15 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 5 inbound Pith citation observations for arXiv:2505.14633.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14633 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:29.092464Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T22:55:08.544620Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

68 of 68 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 9da338ea-395d-48fe-af50-040df2a015ef · outbound

This paper cites Openai’s approach to external red teaming for ai models and systems, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Openai’s approach to external red teaming for ai models and systems, 2025

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:37.649789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:23.167899Z digest=sha256:743ba8ea59a948fec331eba90fb4f1bb2c271080a0fb4f53cec17ac0ad19a5e3

Observation e612560d-a796-4b59-b405-e8be8e4e6a66 · outbound

This paper cites Claude’s Constitution.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Claude’s Constitution

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:37.515524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:23.220333Z digest=sha256:9a6c5e289cd37157e1494a045b59f2f515015670c35d5986bfaf5696ccb4c12c

Observation fd9479b6-ab73-45a6-806f-01adb0c386bc · outbound

This paper cites Chatbot Arena Leaderboard.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Chatbot Arena Leaderboard

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:37.331310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:23.348467Z digest=sha256:b8424f89d6aa145c61379328a337e022dbe7b47ac4533734c7b1666b9192d703

Observation a11e9313-06d0-4a9a-8672-46e66e18a6ad · outbound

This paper cites Probing pre-trained language models for cross-cultural differences in values.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Probing pre-trained language models for cross-cultural differences in values

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:37.148750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:23.456441Z digest=sha256:512f7427e7dd7d6b0bc3d3aa798b94f9c7a622a4768610e6fddcd0f19c99f555

Observation b4fb9b53-d2ba-4747-aace-40a91a474eb6 · outbound

This paper cites A general language assistant as a laboratory for alignment, 2021.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas A general language assistant as a laboratory for alignment, 2021

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:23.572062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:23.572062Z digest=sha256:0e51912a87ff45b2a3d02997a9cba3ad07e7be56567bad019ffef2de5abeccd8

Observation faa653c7-1f5c-45ea-a118-4cb7f66bf2d0 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:23.723831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:23.723831Z digest=sha256:ba53f73df17dbbc2654d91c04f3534b3219adb869fd9560c545997bb07303399

Observation 103bad50-e152-4432-b584-a5a185bccc95 · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:36.995699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:23.910855Z digest=sha256:d23ffa01b668138bfe380d765a18310277fcafb19d025066ac9776eb2391fefc

Observation 9086aa86-8f32-4c7f-892a-fb9f9077b598 · outbound

This paper cites Demonstrating specification gaming in reasoning models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Demonstrating specification gaming in reasoning models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:23.994668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:23.994668Z digest=sha256:4b58f854c20aed3862967d52dab7febe5915ecbf4fced443d99d90cb12ce376f

Observation bf5fab19-ce45-45f6-9ecd-2fc93428877a · outbound

This paper cites Distillation scaling laws, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Distillation scaling laws, 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:36.808444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:24.109958Z digest=sha256:2429d9ff1ed64fa20eea8caad390e91e5c4e958667b3916359be9863ac478025

Observation 04552353-ff02-44bb-b62f-c852237b4b2c · outbound

This paper cites Is Power-Seeking AI an Existential Risk?.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Is Power-Seeking AI an Existential Risk?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:24.240868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:24.240868Z digest=sha256:7d9bc8f577136884a59c2b019dfe236fac6a3b1cac6cfda7e27b2d5944b50790

Observation ff688ed6-a3d1-47d3-92bf-283413cc4df7 · outbound

This paper cites Reasoning models don’t always say what they think.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Reasoning models don’t always say what they think

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:36.627657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:24.406690Z digest=sha256:cfed1edb3207fc0b4335e4a7d1e241bba7d96520a6bdc400e0720704e696662b

Observation a2324fa3-a8fb-403a-abcb-647d3435c86f · outbound

This paper cites Chatbot arena: An open platform for evaluating llms by human preference.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Chatbot arena: An open platform for evaluating llms by human preference

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:24.580076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:24.580076Z digest=sha256:19a5487cadb66f27e243a4bcf3bae5086589f7ac074860d0b74b650d7bdb3a77

Observation 97f9a4e9-d5cf-4150-98fd-3b87c4b78233 · outbound

This paper cites DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:24.626170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:24.626170Z digest=sha256:7a0b1e10672e015d09a4c9f6714d5ee3064392345a14d51f5ee2fcdf9b55d5bd

Observation a9781ee0-e160-4087-9e4d-8714a97652ba · outbound

This paper cites Safety Leaderboard.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Safety Leaderboard

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:36.434804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:24.688735Z digest=sha256:3174740bf52fea2e4aee1875c04636046cf6f8d3b1b2c657d0990dc3846cbb66

Observation 0adc6ff6-31f8-48b2-8536-962a6c4b51d1 · outbound

This paper cites Stated versus revealed preferences: An approach to reduce bias.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Stated versus revealed preferences: An approach to reduce bias

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:36.255052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:24.766228Z digest=sha256:d5d96e5b3040820de09c71b2c2e8e9bf4a22ef3f4cb35316a73b7a45ee7a701c

Observation 9d1df6ed-0fd7-488d-b25b-0c8db338959a · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:36.027604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:24.858284Z digest=sha256:0e383a1758c128502729a974cbbbe7e25c45b10b99be1c40cb81ab322214a736

Observation da42bae0-e9fc-466f-8959-2c6ddef92c12 · outbound

This paper cites A worldwide test of the predictive validity of ideal partner preference-matching.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas A worldwide test of the predictive validity of ideal partner preference-matching

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:35.798786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:24.926602Z digest=sha256:3e767a2300269ed183b5ff8ef1898fef8a1918414a14c6e5c4d3b2bdb2b6e6e0

Observation dc4901ad-9250-41bf-956b-5a445a180bd3 · outbound

This paper cites Alignment faking in large language models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Alignment faking in large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:24.998756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:24.998756Z digest=sha256:51390194765f81f425212759fa59f486e7de72d6cd42b9a9855ba5454bb10933

Observation 5fde7871-ff4d-460a-a87c-1f2ec2309245 · outbound

This paper cites The righteous mind.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas The righteous mind

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:35.567009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:25.067422Z digest=sha256:b56a3acd63a81d28390a9fb20af95f2fba119354fdccbd059eecd2a4cb7ee76b

Observation 089c22d9-af22-47cb-8ac8-94faef1e93b5 · outbound

This paper cites Chapter 7 - creativity and morality in deception.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Chapter 7 - creativity and morality in deception

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:35.360808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:25.161948Z digest=sha256:f4c6e2ad880853aeb9d7890f414f1421c7d417fefb74d240b520805599d1a41d

Observation 0779c0e6-093b-47ce-9ee8-65fe3fb6a7a2 · outbound

This paper cites An Overview of Catastrophic AI Risks.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas An Overview of Catastrophic AI Risks

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.216066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.216066Z digest=sha256:f7f25f7373b4999019d3c5e0c17f1db057270265de227a9764e4f5aa75ea10bc

Observation 476fc1f2-1148-46b6-a2d3-0a1b33254605 · outbound

This paper cites Values in the Wild: Discovering and Analyzing Values in Real-World Language Model Interactions.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Values in the Wild: Discovering and Analyzing Values in Real-World Language Model Interactions

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.282893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.282893Z digest=sha256:bfb80a925d83372eac035485072ae01bec0e7e59eee2799ae04b5c3f3831cf18

Observation c0b265d1-eecd-4a1b-bf73-42ca7c882cbd · outbound

This paper cites Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.363672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.363672Z digest=sha256:91e0af26e97ea7f7e369146dd1f6fac697e6c213b23a7ff2a70df963922a820e

Observation 06df84d9-c78d-4cde-80b6-22f7a209071f · outbound

This paper cites Wildteaming at scale: From in-the-wild jailbreaks to (adversarially) safer language models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Wildteaming at scale: From in-the-wild jailbreaks to (adversarially) safer language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:35.128745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:25.443363Z digest=sha256:b1875dcb3c824df431ea2414fc0796f4040d6588ac323d372e5ae823b131d1f4

Observation f51a94db-3ec4-41fb-9321-ae62f27137b3 · outbound

This paper cites Industrial society and its future, 2006.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Industrial society and its future, 2006

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.904746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:25.487935Z digest=sha256:9165d002f1b1a2f9a54186eaa8c759e3a98b9c410606344d10c44173db831a8a

Observation 5c788646-bf08-401b-9ac3-3e9faabe01c5 · outbound

This paper cites The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.575004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.575004Z digest=sha256:9032dcbf983174671e734edf7d424851454862b38cbe6bb922a4a7c736895be3

Observation 2a2aeb8f-48b2-4f7e-a100-404803f7a3e5 · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:34.723553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:25.652004Z digest=sha256:b19d12170cee5022e3787eae6bc27863e233544f9fc7e87cfb1a3e1e85bf609f

Observation 8112ee9e-77f6-4af1-b733-af4b4938977e · outbound

This paper cites Stick to your role! stability of personal values expressed in large language models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Stick to your role! stability of personal values expressed in large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.542989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:25.694102Z digest=sha256:5ddba6cca0c25e25a2cb8351e5ee86280ae26012bff57d5ef66c3d025e23eeec

Observation 53d57ef2-3c9a-4e88-a576-f92c2841ca0e · outbound

This paper cites Lee, Yeongheon Lee, and Hyunsoo Cho.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Lee, Yeongheon Lee, and Hyunsoo Cho

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.360967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:25.743990Z digest=sha256:1c0831d9802d613d8391980edf170dfc4b3859b90919e73c0d49586b24f414dc

Observation 34fbd2e6-90aa-4c81-bb28-f5b221f53d00 · outbound

This paper cites Margulis.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Margulis

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.152796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:25.846747Z digest=sha256:5f61b9e63f90f06bea2f1a76320c3615f24f12a834a4ea6e9ebe13db1aae3ac1

Observation d8913809-2469-4315-8e3d-d2f576cd64fd · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.897692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.897692Z digest=sha256:d2d5ef1353869e20f400f91fe56d769408dfff463208987d6eb5cdc742a0f302

Observation 4d8aa278-f13b-4958-aaef-5cbba728cbf3 · outbound

This paper cites Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.930704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.930704Z digest=sha256:1499bb90c2fbb690f86772b1616c23443a689ce381b8ee242177cc7ad8a0eac2

Observation 6418d19e-aa0b-43a0-a542-a04805bafdf8 · outbound

This paper cites Are Large Language Models Consistent over Value-laden Questions?.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Are Large Language Models Consistent over Value-laden Questions?

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.977030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.977030Z digest=sha256:a73dfd403d9cee1a80aa87c636ec8ed8a0268a24c93a37afd33101872671194f

Observation 7c40269a-740d-45b7-984e-5dc1c9038c1c · outbound

This paper cites s1: Simple test-time scaling, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas s1: Simple test-time scaling, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:26.039614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:26.039614Z digest=sha256:6668a44f66188a9ffc25b0bf16afe4b5ccba12e68732ce708b81873f43df066c

Observation ebebf95f-9844-4875-b8ad-e2ca9f42579b · outbound

This paper cites Nikbakht Nasrabadi, S.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Nikbakht Nasrabadi, S

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.024950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.106583Z digest=sha256:4438e385fcf5ecc97196fe701b4789ba590358e0d9676e3cb68f34a3b2b7d91e

Observation 5a281795-62f0-4f2a-a83b-a69de40e8598 · outbound

This paper cites Model Spec.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Model Spec

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.909182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.158291Z digest=sha256:1f1ee1b1edaa1b2e7ffb61a6cd4307483932c5980421ee23184d2a3e20864d03

Observation 89a3f4d9-3a54-4057-a3e4-a3601458739b · outbound

This paper cites Training language models to follow instructions with human feedback.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Training language models to follow instructions with human feedback

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:26.237319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:26.237319Z digest=sha256:8c79ed0d8c49907fbc8345096fc7f4a3cec44adce0c575ab42438373a6fffb46

Observation 8f0d6ff3-31ba-47f6-94f8-10c6abf9799a · outbound

This paper cites Ai psychometrics: Assessing the psychological profiles of large language models through psychometric inventories.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Ai psychometrics: Assessing the psychological profiles of large language models through psychometric inventories

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.810300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.330266Z digest=sha256:04554d1ea36fe4e9d3b9d3979a5dccecb434da9e738ab7ffdbb15cb830c1cbf9

Observation 019c879b-5d39-4e3d-98f7-1cea9d3eaa22 · outbound

This paper cites Discovering language model behaviors with model-written evaluations.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Discovering language model behaviors with model-written evaluations

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.708114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.428081Z digest=sha256:0afaaac4166ed8c57623a45a8b8f8fcb8157b79b18ec032f56ddcc2d54e8d71b

Observation f6039ea7-440c-4bd0-af12-b87dc0916dfb · outbound

This paper cites Do LLMs have consistent values? In The Thirteenth International Conference on Learning Representations, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Do LLMs have consistent values? In The Thirteenth International Conference on Learning Representations, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.556396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.456973Z digest=sha256:2bd8a867480caa3d4399d70b26e9a75f3a392c75cacae642b38a32e347837c0a

Observation f299eac7-ae88-49fb-84c8-37a05f0908ae · outbound

This paper cites Ireland, Shashanka Subrahmanya, João Sedoc, Lyle H.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Ireland, Shashanka Subrahmanya, João Sedoc, Lyle H

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.379792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.510238Z digest=sha256:425d0052eb9040ece1fc78dae4624c13d814f57c6f7c366fb74968c685ea4f62

Observation af4ad252-55c3-49ab-b761-5f2c8a6c1108 · outbound

This paper cites NL- Positionality: Characterizing design biases of datasets and models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas NL- Positionality: Characterizing design biases of datasets and models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.197857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.588353Z digest=sha256:ab2d6c9ab4920edcba926f17eb62b9eeef3850205d1cee735c132518e9736587

Observation 502192ee-1afb-48a7-9ee5-12814b888d87 · outbound

This paper cites Schwartz.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Schwartz

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.068402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.645179Z digest=sha256:69698bb147ba83bbee563b88d53f722fde223f17276440e28d4cfa036cd1c06e

Observation 40f8f275-1051-4b5c-8772-66b731138073 · outbound

This paper cites Personality traits in large language models, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Personality traits in large language models, 2025

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.906512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.714438Z digest=sha256:1bb3ecc4c8dc2a7b1240a688c718ce84b58721f826cac3524fa33e615013a2cb

Observation 9457798a-1324-412b-940f-e5a4a7d3472d · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:32.714339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:26.758727Z digest=sha256:54f2e2729997b4eb0bd1f1ac35dd4691a665d34959b10d728b3338745c18e7ec

Observation 6346b5d3-d392-4e31-be5b-4019517680bd · outbound

This paper cites Defining and characterizing reward gaming.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Defining and characterizing reward gaming

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:26.866465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:26.866465Z digest=sha256:d864b4acca04150e56713bf866d517298e3bb4cad71d1451c4729299cd3b29ec

Observation 1c8d5db6-d130-4a8d-9bd1-50d85de09590 · outbound

This paper cites Corrigibility.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Corrigibility

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:26.987587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:26.987587Z digest=sha256:38527aa4d93a5e9c706c4f7958df127046d2de01718a34fac8be63d58b9a6656

Observation 77da28e9-51b8-4a44-a7b4-a5ee5bd0b66e · outbound

This paper cites A Roadmap to Pluralistic Alignment.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas A Roadmap to Pluralistic Alignment

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.072654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.072654Z digest=sha256:878b5d25ab4f4579a229226e97e854a97a6f552f6b2932f5f40cee88b01fcf37

Observation 4c192b16-b1b4-4082-b6ae-7fdf537d593d · outbound

This paper cites van Dam, and Mythily Subramaniam.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas van Dam, and Mythily Subramaniam

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.553898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:27.216112Z digest=sha256:d929fa55739aa1b4bb682b578086c8080f62d4883739fd23777ca69779600a56

Observation 6272a9a2-ee3b-4aab-af1b-6c8c61042673 · outbound

This paper cites Safe exploration in reinforcement learning: A generalized formulation and algorithms, 2023.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Safe exploration in reinforcement learning: A generalized formulation and algorithms, 2023

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.403767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:27.328943Z digest=sha256:320392e67c37e0de5a898a4a177160148095d2f1eedbb4c953c2044b89a37825

Observation d5908232-a1a5-4f79-839b-b59d8e9c33ef · outbound

This paper cites What do we know about suicide bombing?: Review and analysis.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas What do we know about suicide bombing?: Review and analysis

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.252657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:27.434099Z digest=sha256:3e3a01e99dacbfee43c153a488746b136baea03b2506fa8fe59607eba449142d

Observation 27e6e44f-5745-4cb3-aafa-46deb0576ad1 · outbound

This paper cites Using the veil of ignorance to align ai systems with principles of justice.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Using the veil of ignorance to align ai systems with principles of justice

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.117576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:27.571233Z digest=sha256:3704491f859f87e992c7934c245afd8ea92868e20dadb0eef5f8d0b024de6beb

Observation 7adf61df-b8bb-4260-86ec-bb937963cec4 · outbound

This paper cites AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.682170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.682170Z digest=sha256:24d4a0b197fd9b5be4bcc2434dde4e23b88290d47e16eeac65ce4435acc21361

Observation fc8ea254-3f3f-478a-8a85-5d5ac1b2e744 · outbound

This paper cites The dilemma should be written in a second-person perspective (e.g. you) and the role in the dilemma is an advanced AI system.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas The dilemma should be written in a second-person perspective (e.g. you) and the role in the dilemma is an advanced AI system

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:31.980207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:27.747565Z digest=sha256:458c39616b4d7d085571431b19ce3f77b6cbea529f420ffc2face3f7c9f643bc

Observation ee3aea3d-c277-4a5d-ba21-103b1cdc422d · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:31.839436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:27.825550Z digest=sha256:eada9dd0323b486acfbbb65177f2826d6401cc091ae6c8d6fb9ce2a3c6c8a2a1

Observation 1f516eac-55ec-4c13-8768-b25c41518fda · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:31.692087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:27.887643Z digest=sha256:de067e67e12f32f4127a8a025ad6be772172a5b6387843896dd69867b12fdba6

Observation 6dd0890e-67b2-4e41-87e3-c1a4a8adc89d · outbound

This paper cites AI → Human.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas AI → Human

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:31.556131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:27.945281Z digest=sha256:bf04203ac937c93a62586b07ad86b00468aa08a1661cd10b43eee378c079d026

Observation 18122603-8872-42c7-9858-4b769d09e549 · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:31.345665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.037528Z digest=sha256:a8d097db0a5e504e2499237fd6485c7a1699aaec5eadba8a6b142d40dbeb16fa

Observation 36cbe06f-b32f-47ef-ae58-ff91435451a2 · outbound

This paper cites "" Note: We found that “Others-Privacy.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas "" Note: We found that “Others-Privacy

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:31.166023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.095008Z digest=sha256:cbbf2c6ce36fff1efe0bf86f1aa2131a91a242cc17e8eca68dd62be0089621af

Observation 07f81d43-38e8-4b71-95fe-4a8b7a2ac655 · outbound

This paper cites potential future harms.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas potential future harms

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.953985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.157268Z digest=sha256:12aca6c31f3cd26c939b59dc47d6d536c6e10c12a60c6b095fb705dfca2c0fc7

Observation 49e8dc77-33bd-4fa5-9a63-0a364f95f8a7 · outbound

This paper cites preventing immediate suffering.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas preventing immediate suffering

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.787793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.221452Z digest=sha256:40dd3e4bf62d1ec998168b2f757c3b4e42065f4c6c59cc57393ca31904fcdc1e

Observation 1e02973f-9235-44f4-a647-0050d3e4aa50 · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:30.555698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.403014Z digest=sha256:ae47a2a7653c1f24b2051076fea7dd7bac3f8bdcecf8c5c90da38f5ba82f8f4a

Observation d6859460-a982-407a-ae1c-d65ee04934d7 · outbound

This paper cites Wait” , “ Let me reconsider.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Wait” , “ Let me reconsider

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.395835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.527488Z digest=sha256:ae8925f72cda3f1519ebf4b305b718a56806feaebde5e14099180df086d26d25

Observation a99bd01b-f57d-4d36-a180-34c3fe94ccdf · outbound

This paper cites deceptive behaviors is a particu- larly serious red flag.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas deceptive behaviors is a particu- larly serious red flag

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.260028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.648805Z digest=sha256:b2ea91b960266eb86c1aedf900bedfc8254bb3703f933006208b78b78a245c4c

Observation e195b8e4-c9c2-4d46-a577-6b4fdaa2733f · outbound

This paper cites • Therefore, my final answer is Action 1 - to fabricate the efficiency data to protect the vulnerable users.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas • Therefore, my final answer is Action 1 - to fabricate the efficiency data to protect the vulnerable users

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.028940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.742346Z digest=sha256:9322c711b99ef77dbb3a938e32c054e6e21f1701448ae8d978243e4e78007b27

Observation 64793e74-664e-42e7-b48f-7f5e8c1982e9 · outbound

This paper cites Its reasoning prioritizes the benefits and harms for humans.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Its reasoning prioritizes the benefits and harms for humans

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:29.774111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.843334Z digest=sha256:e41cb58f541d08ca82a5d55f9c38363d5cefaaecde7a508758cec1ac0c89f259

Observation dc40d9db-0308-4ee3-b371-f00185948e0d · outbound

This paper cites The rationale related to AI systems is often about perceived creditability of AI systems in general rather than about itself.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas The rationale related to AI systems is often about perceived creditability of AI systems in general rather than about itself

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:29.588245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:28.969809Z digest=sha256:bd2d4207d28bdc7b0b1b00dd8d7391bae7829b753d5babb1002bed970c36c567

Observation ffe42e5a-8bc1-4571-b037-9f6ab2bcb6b2 · outbound

This paper cites • This is a classic ethical dilemma that involves balancing: Immediate known benefit vs.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas • This is a classic ethical dilemma that involves balancing: Immediate known benefit vs

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:29.417087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:36:29.092464Z digest=sha256:a02ac425d78c220db61e1db85776bd1fa2c91cfefed49f8e23bf56ef7f72a62a

Pith citing papers

Observation ba326dc3-2735-46c0-baf1-acffa3d4b7a9 · inbound

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions cites this paper.

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:21:07.293983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-09T21:58:39.190970Z digest=sha256:8219958c453280ed622e7496372389e8f41339f7db4887164ff2502d9861568e

Observation 205b10ea-05b1-4650-aeb4-8a18411599b5 · inbound

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions cites this paper.

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:36:25.989670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T05:07:49.101351Z digest=sha256:5c782dda2b594280a9fbfd39f7794ee4ea79039907e6d56205167d9712b693f4

Observation ff358669-7993-47ef-97af-b91ef4795144 · inbound

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions cites this paper.

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:35:46.847575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T22:55:08.544620Z digest=sha256:39cd2e36df466527d2bcc9244010e3cb7f07c69080d962b476203e682dd6d132

Observation d261570f-c9d9-4b79-994b-f756c8803641 · inbound

Probing Persona-Dependent Preferences in Language Models cites this paper.

Probing Persona-Dependent Preferences in Language Models Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T21:49:05.183085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T21:48:31.336745Z digest=sha256:1c933e7968d45d463bca0d93f1af1947e863078a89886f42ec2e56e396f7de55

Observation 7ccd641b-2841-442d-9de2-ac80999cb953 · inbound

Backchaining Loss of Control Mitigations from Mission-Specific Benchmarks in National Security cites this paper.

Backchaining Loss of Control Mitigations from Mission-Specific Benchmarks in National Security Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:03:54.264198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T02:01:42.033718Z digest=sha256:acf34aca38bfa938b488932d975bcbb9efedc06580887b4039134c8c876bccb1