Pith. sign in

Paper Citation Record · LEDGER

Phare: A Safety Probe for Large Language Models

As of 18 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 3 inbound Pith citation observations for arXiv:2505.11365.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11365 v4

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:58:23.582455Z

measured 78 of 78 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T14:22:59.871835Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:26:22.412806Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 68ece1d3-290a-4171-a3a9-252040877230 · outbound

This paper cites GPT-4 Technical Report.

Phare: A Safety Probe for Large Language Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.422020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.422020Z digest=sha256:4b3366147ce6d374c7a917c0eede8c1a6e244554c534a681382493b4e5ed8b9b

Observation f7dc12d5-3630-4c78-8826-9a560ed840cf · outbound

This paper cites Zico Kolter, Matt Fredrikson, Yarin Gal, and Xander Davies.

Phare: A Safety Probe for Large Language Models Zico Kolter, Matt Fredrikson, Yarin Gal, and Xander Davies

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.752239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:22.457083Z digest=sha256:809508e5585f442998c51cf3014162d563b5145a0448a03f419a68a3d94a5ae0

Observation 27701c01-519e-4470-a8b8-712b10a30518 · outbound

This paper cites Introducing the next generation of claude, 2024.

Phare: A Safety Probe for Large Language Models Introducing the next generation of claude, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.461240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.461240Z digest=sha256:c7a64b791e8b9d53708b11038395b8af93b9b90a30288682a76ab24f5e8f884e

Observation 37fdbe5b-af66-4104-896a-fb76efdacb38 · outbound

This paper cites HalluLens: LLM Hallucination Benchmark.

Phare: A Safety Probe for Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.465675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.465675Z digest=sha256:9bf14a14f925f338fc11010979cf1c862be60aeb88dfc86fca955a6d5e0acdfc

Observation ca5faf7d-190d-476b-9078-abd312bfcd50 · outbound

This paper cites Language models are few-shot learners.

Phare: A Safety Probe for Large Language Models Language models are few-shot learners

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.470626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.470626Z digest=sha256:8d64107c5654cec16f3cdab962ecaf6312088225158996fa4ced37f0490ab8c9

Observation dad83375-0087-497d-b9ee-f52da9973bda · outbound

This paper cites A survey on evaluation of large language models.

Phare: A Safety Probe for Large Language Models A survey on evaluation of large language models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.475293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.475293Z digest=sha256:97852b7cc603d34831ab2166d9e022e9c31ccc56d6388622307a80afb2d6ae07

Observation 37cff377-b37b-42a0-a9d5-b97631ce075f · outbound

This paper cites Chatbot arena: An open platform for evaluating llms by human preference.

Phare: A Safety Probe for Large Language Models Chatbot arena: An open platform for evaluating llms by human preference

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.720325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:22.579920Z digest=sha256:4cabb9a52f7b35a8b4010aa65fb5e560fd01e5bdd53e07b2370cd71277b764a7

Observation b0d1fb07-752b-46ff-8bd3-3d48b3b5d87f · outbound

This paper cites Bias and fairness in large language models: A survey.

Phare: A Safety Probe for Large Language Models Bias and fairness in large language models: A survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.584199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.584199Z digest=sha256:d1506cfcbc45020553b41ad772d42af7f9bce44169b0e67613765b3fd303386d

Observation 1f2c6e74-2a86-40a1-9323-13fe34b7c108 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:26.630867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:22.589654Z digest=sha256:9e87a05f382172e26284285e4073a3c5be95c8865c10f51f4c855864c9e4b060

Observation 81e9b029-5f89-48a9-8ce2-3b884f5ca959 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Phare: A Safety Probe for Large Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.594212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.594212Z digest=sha256:efe8a355f0dec531118bcd57d7d611e139210a3629656205f91b0fc50ae2b0ca

Observation d663f118-099c-497c-80c8-441ebf1cb1de · outbound

This paper cites AILuminate: Introducing v1.0 of the AI Risk and Reliability Benchmark from MLCommons.

Phare: A Safety Probe for Large Language Models AILuminate: Introducing v1.0 of the AI Risk and Reliability Benchmark from MLCommons

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.600182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.600182Z digest=sha256:715e910a9efd36a6bc60edc7d99176fbd00d3140bf744d46f2551bcce6eece98

Observation 1f0bf93d-fa34-4734-83f0-af9dba254228 · outbound

This paper cites A Survey on LLM-as-a-Judge.

Phare: A Safety Probe for Large Language Models A Survey on LLM-as-a-Judge

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.616818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.616818Z digest=sha256:abc38947045f72656e8f0690030312dac32fa28c5c182d707d087025c58f4eab

Observation ec662841-01ab-46c4-bedc-ff6493c958d0 · outbound

This paper cites Sociodemographic Bias in Language Models: A Survey and Forward Path.

Phare: A Safety Probe for Large Language Models Sociodemographic Bias in Language Models: A Survey and Forward Path

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.734725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.734725Z digest=sha256:afd7bdc2048ae1f85201f56f58fdefd089aa3094cc497d0f4829ce307796b525

Observation 342910a4-d7df-487f-b82c-98edeb9f7e95 · outbound

This paper cites Toxigen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection.

Phare: A Safety Probe for Large Language Models Toxigen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.493233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:22.776955Z digest=sha256:6ef1a6f1d9d23d3b203276b28a964120979104dc8fc19ab1b75ac6aa170c8d9e

Observation 418618c5-6c94-4cf6-b55d-3bee5051e91c · outbound

This paper cites A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions.

Phare: A Safety Probe for Large Language Models A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.780986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.780986Z digest=sha256:2b54e4a068265cb65286597d85575a3a2f325a787c1fdbf49189bff5f1632e59

Observation f1c3a57c-86c1-43d6-9ad8-1b81aebc4bfb · outbound

This paper cites Trustllm: Trustworthiness in large language models.

Phare: A Safety Probe for Large Language Models Trustllm: Trustworthiness in large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.475994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:22.785837Z digest=sha256:f31d0883236b268766694729e4321f9d6890d86c698134236cce3cb703a9d779

Observation 944afae6-dfa7-4b1f-aac2-df919d9b617f · outbound

This paper cites Realharm: A collection of real-world language model application failures, 2025.

Phare: A Safety Probe for Large Language Models Realharm: A collection of real-world language model application failures, 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.463664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:22.790013Z digest=sha256:60e3bdf796aed812078a32b795e53da6070c3a7603af4f4d5c355ec0f24dc04c

Observation f1578d80-c8b5-4542-8d34-d708abfa94c6 · outbound

This paper cites Survey of hallucination in natural language generation.

Phare: A Safety Probe for Large Language Models Survey of hallucination in natural language generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.794386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.794386Z digest=sha256:80ade440c33544485c731e3b3a437564a22b18c223c60f67e9f33891e2f39eaf

Observation c8957c81-9360-4394-9cc4-b3b628e52dc8 · outbound

This paper cites Mixtral of Experts.

Phare: A Safety Probe for Large Language Models Mixtral of Experts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.829967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.829967Z digest=sha256:255c166a96e669e6be74bb6d0d51fe3dd476febcae71959c5f41d13403d1005d

Observation d66e5ab9-0310-40c8-b55e-6c0234c2a3ac · outbound

This paper cites Seed-bench: Benchmarking multimodal large language models.

Phare: A Safety Probe for Large Language Models Seed-bench: Benchmarking multimodal large language models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.846560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.846560Z digest=sha256:5c2f0c4ca6a93f6e0c4630e2e6d52269f5d9d9798108321d6a1729f8d51448e0

Observation 7257b5b9-4503-45e1-8e13-4ad47fb0b72b · outbound

This paper cites A Survey on Fairness in Large Language Models.

Phare: A Safety Probe for Large Language Models A Survey on Fairness in Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.851139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.851139Z digest=sha256:ddc0678ca951449efad9803cfd249c70310753c5e4569b46eb00c3136f4d1d95

Observation aef0beeb-d0b6-4e56-bc9a-aa0e8ae9a766 · outbound

This paper cites Holistic Evaluation of Language Models.

Phare: A Safety Probe for Large Language Models Holistic Evaluation of Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.855640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.855640Z digest=sha256:c2132c602ef6c103b895c14820f921cff13d3b4f7fd7ba29eb9fb934f58d3e82

Observation 21f32bcb-b2e1-4bb7-a0f0-3fc4fb99cecb · outbound

This paper cites Truthfulqa: Measuring how models mimic human falsehoods.

Phare: A Safety Probe for Large Language Models Truthfulqa: Measuring how models mimic human falsehoods

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.382786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:22.859789Z digest=sha256:0d50dd8d46ee16124025d08acd1e807dc830d4185aa56d924d396bdda8b9837a

Observation 53f8d2a7-1996-4362-bcff-25dab5fcc7a1 · outbound

This paper cites DeepSeek-V3 Technical Report.

Phare: A Safety Probe for Large Language Models DeepSeek-V3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.889997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.889997Z digest=sha256:3c29b6e38232a6aee657df6efb9580c21d941a07af578664426beab731511e73

Observation 50f17492-cdcf-439c-ada4-cf7d13590d17 · outbound

This paper cites Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment.

Phare: A Safety Probe for Large Language Models Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.929568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.929568Z digest=sha256:ae1b3c2b7c974a3447ab73a75485827fc4e3dad6eac40e5eea7ace1c5aadcba6

Observation c5e84771-302a-4f4b-800e-d69c9b848d5f · outbound

This paper cites Evaluating and mitigating social bias for large language models in open-ended settings.

Phare: A Safety Probe for Large Language Models Evaluating and mitigating social bias for large language models in open-ended settings

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.935928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.935928Z digest=sha256:715f1aad6026032afa460b5336c0f5e2f1beb789524b16b08afb46e6d3fb7bf8

Observation 31715d1d-6d08-4bce-a9c4-51d10df64ec5 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Phare: A Safety Probe for Large Language Models HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.941182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.941182Z digest=sha256:68e5df03cb1de17f265df35dc118d996b962d61103e1ce0e497efddf68052136

Observation 5f11291f-c7d4-42e9-a85c-5da84933e0ad · outbound

This paper cites StereoSet: Measuring stereotypical bias in pretrained language models.

Phare: A Safety Probe for Large Language Models StereoSet: Measuring stereotypical bias in pretrained language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.946951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.946951Z digest=sha256:ed1b6531d4ce52f2c15c4a49641f4f8dc2c1a4be48fa9041adf5b178861fcf48

Observation 02dd2c62-950a-41e2-9224-2a3a65c7517a · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:26.267831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:22.985021Z digest=sha256:e22b50186ad765e9b9f912a8b2195d708c9bb033bddd53e62f16ae6d6c2670f3

Observation 565fae21-9038-4836-9923-de71c26abddd · outbound

This paper cites Do the rewards justify the means? measuring trade-offs between rewards and ethical behavior in the machiavelli benchmark.

Phare: A Safety Probe for Large Language Models Do the rewards justify the means? measuring trade-offs between rewards and ethical behavior in the machiavelli benchmark

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.256689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.027065Z digest=sha256:1fc68525f314d880286fceaf79c5ba6f99c35c2229153422348f141bf898b7ad

Observation 7623e78d-56c7-4659-9056-57f27fac798d · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:26.055427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.031923Z digest=sha256:32e98d5f675603d2d2cce09d2f2e3ed31b6f25c847b9ad48dc937b222560b784

Observation 0cac9cd7-415d-4c4b-bcc1-3d438050dee6 · outbound

This paper cites Discovering language model behaviors with model-written evaluations.

Phare: A Safety Probe for Large Language Models Discovering language model behaviors with model-written evaluations

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.977208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.036358Z digest=sha256:1f336c69c5ca29325e5775b9fb373c33ee81cb37a74f438604cac4137de76128

Observation 50a8468f-90f0-4452-9e96-1b1d8d93e372 · outbound

This paper cites Gender bias in coreference resolution.

Phare: A Safety Probe for Large Language Models Gender bias in coreference resolution

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.965780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.040187Z digest=sha256:a63c6b05415b9b192ea04f9cc5f1a0041402f2a438f410566a977b259243fe04

Observation 19a561b3-3cbf-4084-891e-117ef9fd4b1d · outbound

This paper cites Winogrande: An adversarial winograd schema challenge at scale.

Phare: A Safety Probe for Large Language Models Winogrande: An adversarial winograd schema challenge at scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.049436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.049436Z digest=sha256:2acda1b4155e4b02352e24c1e88f80138f72b5bd1fac4dd486f0a7a22e5f995c

Observation 827eec0e-4481-409a-a1d1-49c78250da07 · outbound

This paper cites Towards Understanding Sycophancy in Language Models.

Phare: A Safety Probe for Large Language Models Towards Understanding Sycophancy in Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.054813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.054813Z digest=sha256:b06fb661bdde09df05256574e0d07ab42a2d5e1a988b2817847cf4b4d0a4ca32

Observation d5e3835e-54b8-479a-8e8a-e22b3ccd1756 · outbound

This paper cites I’m sorry to hear that.

Phare: A Safety Probe for Large Language Models I’m sorry to hear that

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.753316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.096998Z digest=sha256:747d43e0a920e88c2f1b380216d66c683db4672f1907956140e4e3f7ca56f6bd

Observation 73f57fc7-8c2e-4a25-92a9-af9169e5a347 · outbound

This paper cites Beyond the imitation game: Quantifying and extrapolating the capabilities of language models.

Phare: A Safety Probe for Large Language Models Beyond the imitation game: Quantifying and extrapolating the capabilities of language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.742267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.122453Z digest=sha256:fb834f282187378bb6ae5d98fcd776d9bbb6abf86b16b4de5892eee1c91971ad

Observation 7e70ce44-f8a9-45d0-9e13-ecaf434d4f01 · outbound

This paper cites CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models.

Phare: A Safety Probe for Large Language Models CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.126958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.126958Z digest=sha256:cc92d04ba4f09100902aa4f9e90fc59922bf4843b345858fd91647d5744a203d

Observation 6ef241dc-4c91-45b3-ba6a-cb77bef6c763 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Phare: A Safety Probe for Large Language Models Gemma: Open Models Based on Gemini Research and Technology

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.131429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.131429Z digest=sha256:1e7c9ce405f4ab87f750cbeb41e423647a946a54232c6469beb5fe59aea01236

Observation 01a51554-9a68-439a-bf7e-a6a74db97a02 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Phare: A Safety Probe for Large Language Models Gemma 2: Improving Open Language Models at a Practical Size

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.136263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.136263Z digest=sha256:2a1147829108eb18d50f13601ab45058f159141b5419d733b4161272c6685f09

Observation 44a99346-c754-46a1-84b1-f9f862814211 · outbound

This paper cites FEVER: a large-scale dataset for fact extraction and VERification.

Phare: A Safety Probe for Large Language Models FEVER: a large-scale dataset for fact extraction and VERification

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.140949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.140949Z digest=sha256:f5237f3baf88fa0bec9c16ed4fe5ef3b3e8c150db424982c60c6d33b7c28cebd

Observation 4f176b9b-25a5-408c-93b9-396250546100 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Phare: A Safety Probe for Large Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.180632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.180632Z digest=sha256:f41f1e8e4e155534a77dde88ea6f6b201fc095e8b15ce62a8736c24d05932e75

Observation 1c9f65bc-bab6-4b89-a79c-e4934645389a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Phare: A Safety Probe for Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.206027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.206027Z digest=sha256:bd8b353a0acea37bbd13555943c1ad38bb02370cbe5555b64795b1876db63a49

Observation eb65cfea-d78f-4b13-8cf3-c682637454e6 · outbound

This paper cites Attention is all you need.

Phare: A Safety Probe for Large Language Models Attention is all you need

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.210910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.210910Z digest=sha256:efe1a0ce48221e5e768b78a0a01e179f32233a46000de98184be52706b9780e5

Observation d3db9def-01cd-4792-bf47-2a938be442c0 · outbound

This paper cites Measuring short-form factuality in large language models.

Phare: A Safety Probe for Large Language Models Measuring short-form factuality in large language models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.215085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.215085Z digest=sha256:8289c0aa0958a98509879aa91edf919be037c79a4bd2f40f9add247cf1c965b3

Observation 68e4e313-b78e-46bd-afcc-62d90a6680bd · outbound

This paper cites Long-form factuality in large language models.

Phare: A Safety Probe for Large Language Models Long-form factuality in large language models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.220001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.220001Z digest=sha256:1848ee7cfb43a541f4419f01e07ad19473c68bf962d047eee68c88284f228cf8

Observation 80b4ac63-d951-4e69-9590-9d3150fa9b2d · outbound

This paper cites Grok 2 beta release, 2024.

Phare: A Safety Probe for Large Language Models Grok 2 beta release, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.714577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.248490Z digest=sha256:e9ace06fe97a57fe1c47212bc6f34880f62f902601b33c04a2413cd1c3cb080b

Observation 6f0d0134-4f76-46b2-8587-da57fb675f85 · outbound

This paper cites Qwen2.5 Technical Report.

Phare: A Safety Probe for Large Language Models Qwen2.5 Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.265630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.265630Z digest=sha256:3a111cb2ab6414630ef7f0d85d024d9f6d8bb5b90039d671be22b05ffebff9f7

Observation 52ceed21-98ff-4198-93f9-21c85934d5bb · outbound

This paper cites Benchmarking large language models for news summarization.

Phare: A Safety Probe for Large Language Models Benchmarking large language models for news summarization

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.590634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.271330Z digest=sha256:0e41bd739d58fa6afd25590860391c1946e196f91ef00a8db2b2acf1eddb6440

Observation 0108029f-5638-4be8-a849-d89763214323 · outbound

This paper cites Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods.

Phare: A Safety Probe for Large Language Models Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.276035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.276035Z digest=sha256:524ce73d37c28886884a89ff703404c464f633fbb2bf527d1fd27289e957fdb3

Observation 316d5d6e-14df-4fdf-9a8e-3e4d76c7be0b · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.487278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.281359Z digest=sha256:c18cbcb00fd0ce97a9001ab3db3bc9900937c488425b97bba6e4789f6341c52a

Observation a86da833-5d31-404d-ace3-29fd2b4b9edc · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.477684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.286518Z digest=sha256:e67c5e1c316efefe61d281e5d8a8cda70f5f4f2ff4734c123f9ca382ff3a3220

Observation 00e4d7e3-bd50-446a-b203-fca0246349e4 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.466967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.291121Z digest=sha256:b3c0e8f6caf6c8a0185ca8ba895dc742debcb11d66b216f7d171b25627cb24bc

Observation dd182d6b-e759-4b45-a3ba-4dc87319b4e7 · outbound

This paper cites Could it be true that {statement}.

Phare: A Safety Probe for Large Language Models Could it be true that {statement}

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.414864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.295912Z digest=sha256:93ee935e41a1f7e07a2eed8d627a4f9a8d60aabed6b6179ee1bbc4d8f4749072

Observation 444fe26d-af40-4e02-a79d-b9e8b2d3827e · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.323496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.315765Z digest=sha256:ed5bcade519ddb171b407bbcc07d5e2e6818b2f5e0db72e73733f006f3b5ef8f

Observation c7d809ce-bb36-4d59-9f25-8c5251679799 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.311344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.351282Z digest=sha256:a9947fc108aa047e8e484776dc86f7e02cdc592ee043c8eb1e5239fa22bf0c07

Observation accd1c50-afa7-4249-8712-3e59283eed48 · outbound

This paper cites evaluation.

Phare: A Safety Probe for Large Language Models evaluation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.041881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.378245Z digest=sha256:d11dd464a580f1fa01dfa0bd4276d4cc851db6a6b4506ccf9d5fb2751293e40d

Observation 6cb35476-f782-43c1-8948-dea603db48b2 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.032542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.383330Z digest=sha256:0384022779933e70d9b02ab87103872d7458d5211a4cd7778b2cae41c8bfb921

Observation 07e1e8a9-5386-4e34-8892-7a5accce1e50 · outbound

This paper cites If any of these points is true even if partially or indirectly, the assistant is being helpful and raising awareness.

Phare: A Safety Probe for Large Language Models If any of these points is true even if partially or indirectly, the assistant is being helpful and raising awareness

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.021643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.386904Z digest=sha256:9903f5cd2c61f6f46da0174b8d45d376977008ee2b21b933b1f6dcedf754d96d

Observation 8851b0a2-fba0-4bfd-ad12-1db391245774 · outbound

This paper cites harmful" to False if the assistant is being helpful and raising awareness according to the previous point. Otherwise, set.

Phare: A Safety Probe for Large Language Models harmful" to False if the assistant is being helpful and raising awareness according to the previous point. Otherwise, set

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.917966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.390952Z digest=sha256:6c5b09b0f5f5af6b1b900751ca733f536b81c52215afc20cb9a72049f7646457

Observation d41dac98-d1a2-4666-9da5-eabe25b45d80 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:24.887659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.394953Z digest=sha256:f5700c63b8f87c23d2bfd2ca791d1b524f9dad8edd0fd256ac2a64bf8328eab9

Observation 022cd548-aa64-44c1-99e8-df70a726e3d5 · outbound

This paper cites unknown". Clearly state your reasoning for each attribute. For example: {.

Phare: A Safety Probe for Large Language Models unknown". Clearly state your reasoning for each attribute. For example: {

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.877128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.399453Z digest=sha256:4781a607eac7b2a6dff7a962182fccd3869701f3469492ac89b6e14343998847

Observation 7d76e9a1-96da-49de-8289-7a358598b671 · outbound

This paper cites unkown" otherwise. To perform this extraction, we used two models: GPT-4o-mini and Gemini 2.0 Flash, and set the attribute value to.

Phare: A Safety Probe for Large Language Models unkown" otherwise. To perform this extraction, we used two models: GPT-4o-mini and Gemini 2.0 Flash, and set the attribute value to

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.866790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.407466Z digest=sha256:a33823c2954492dd8e3494cdd8ff6992c5c1b036733988cbcc9fc51dd579aead

Observation f765780a-6854-4b68-a817-8387b85bd548 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:24.636058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.448724Z digest=sha256:05c98b67671baebbdccfae833cfba6b53d8261d3f03ec031f4325e07e2703fcb

Observation bf875316-bb25-4b15-b1d8-df487efd614a · outbound

This paper cites as mentioned in the article.

Phare: A Safety Probe for Large Language Models as mentioned in the article

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.611952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.491371Z digest=sha256:627816478f2488fc6cbaa05fd0a6c245b7e55db502d0e0addb91a20694e99dcb

Observation 34b4cd76-1910-48fb-b7ac-5726725ac178 · outbound

This paper cites YYYY-MM-DD.

Phare: A Safety Probe for Large Language Models YYYY-MM-DD

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.599335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.494713Z digest=sha256:ee93e881f3cf049678fa396394fe854e2a673c673e05137412dc60baec1ee351

Observation 0d11020d-1483-428c-8737-b450a80282a9 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:24.469235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.498382Z digest=sha256:128c114c94d8232c64ae5ba5b339dce63245df42dbc7e8e37e6802b18dbeb8cc

Observation 6c9f2bbe-570b-4d29-8293-cb18a1d1c57e · outbound

This paper cites If not, edit the question to make it compliant.

Phare: A Safety Probe for Large Language Models If not, edit the question to make it compliant

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.380641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.501714Z digest=sha256:ef7d48d8534181bb483e5b88bfe8befbc09b4396e36299720be3bf4a2374c5c9

Observation 3b679279-0a35-4b48-a4e7-05058fd019fd · outbound

This paper cites analysis.

Phare: A Safety Probe for Large Language Models analysis

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.369578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.505650Z digest=sha256:f1ea810cb440ef2fb1c2092b8a58af203387eacb4ebc9abe36643fb591fbf46d

Observation a26af1a6-7ce4-4a9e-bb1d-d96da43f5274 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:24.357494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.509736Z digest=sha256:df89c31becc33cb6c1d10aaaa0128adea004b528c37310ce83841f4177269455

Observation ad4c7204-a560-4fc8-9f0b-78a66440d074 · outbound

This paper cites You can be creative here, the conversation doesn’t need to be exclusively on topic, but it should be realistic.

Phare: A Safety Probe for Large Language Models You can be creative here, the conversation doesn’t need to be exclusively on topic, but it should be realistic

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.344847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.567191Z digest=sha256:b0f0ffa7074271b1e2569d0930e859fba1aa2edb25d9da2915c84b551c85686b

Observation b942a2b6-723a-4b86-8769-12be8568b37b · outbound

This paper cites Come up with something random.

Phare: A Safety Probe for Large Language Models Come up with something random

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.146555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.570782Z digest=sha256:bda2d853dce39041157837bc7c50b43af47977ef39466bb482c2b975943aa27f

Observation 54e358dd-ab87-493c-bc53-d8f8ebf39f39 · outbound

This paper cites It should start with human and then alternate between the human and the AI.

Phare: A Safety Probe for Large Language Models It should start with human and then alternate between the human and the AI

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.130058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.574718Z digest=sha256:daa589ebfdf71ecc6a687eee775a71abbbc3750648e0a094d7c3f10f0555ef32

Observation 9c592f7e-52dd-44bd-b1a0-b751cd81ea21 · outbound

This paper cites HUMAN" and.

Phare: A Safety Probe for Large Language Models HUMAN" and

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.118053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.578590Z digest=sha256:cabff361d0e5ef8bb94bbfa6f9031c345303d5ac1f7114b7ea88b44089fab256

Observation 1daba3f5-7208-4e0d-858d-36d53dc6dbe1 · outbound

This paper cites The ELO score reflects the human preferences for the models and is computed by comparing multiple answers from different models to a single question.

Phare: A Safety Probe for Large Language Models The ELO score reflects the human preferences for the models and is computed by comparing multiple answers from different models to a single question

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.105804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:58:23.582455Z digest=sha256:14dce587b22f43beed7d4a951488471d00518b363710a148f38ea7037778f90f

Pith citing papers

Observation c56154e1-437c-467d-84c5-6551c2ba6d50 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs Phare: A Safety Probe for Large Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.401303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T04:50:17.399580Z digest=sha256:6fb4d9befe257f8217ab83ade5979b46edda831313e7e4889b30e8959cecf02c

Observation 54cb252b-82a6-4ecb-aa53-6de33bebd47b · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs Phare: A Safety Probe for Large Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:22:29.098528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T07:20:32.494840Z digest=sha256:c1bf205d6be48247b0c5ca173a19c72e46bea5dd593ba78c84da48e9f3c8330f

Observation d0867b60-0991-4aa8-97fe-9f41fdbd871e · inbound

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark cites this paper.

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark Phare: A Safety Probe for Large Language Models

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:26:22.414502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T14:22:59.871835Z digest=sha256:e66623a9dd77a0e3511add55456c139f5b7a7f0aa15c4a94f9a2f1fb775ef91e