Pith. sign in

Paper Citation Record · LEDGER

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors

As of 19 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 4 inbound Pith citation observations for arXiv:2505.14300.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14300 v2

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:33.763079Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T11:56:07.082215Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:37:37.332332Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact2
  • verified fuzzy8
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9cfadd81-d068-4024-a70c-ecc59ea4ab69 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:28.543259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:28.543259Z digest=sha256:84892beabeb3936e5aaf7ce613a3835b83fc88450a1993333abd01b65716a4e9

Observation 5149911b-abee-4c46-9cbc-ff6d294fcf87 · outbound

This paper cites Obfuscated Activations Bypass LLM Latent-Space Defenses.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Obfuscated Activations Bypass LLM Latent-Space Defenses

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.275006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.275006Z digest=sha256:d000a0e5ce594a69caefd2b5fbcb933ff016098169166d661d875f22020614f5

Observation 06bb95d0-789c-442c-a834-1485bcff7de0 · outbound

This paper cites Towards evaluations-based safety cases for AI scheming.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Towards evaluations-based safety cases for AI scheming

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.414746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.414746Z digest=sha256:e3a7af0e79863e888f8f69f2ce369e4c427ae830abefa742136718952a270df0

Observation d161dd7e-dea4-4a28-9099-57735193d468 · outbound

This paper cites Autoencoders.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Autoencoders

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.498720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.498720Z digest=sha256:20097826dec636b78e18c1ca49fa1b7ac092274f8b3da8310395f0e933a60f8c

Observation 1a98e376-f4db-4820-a492-b0f9960b93bd · outbound

This paper cites Taken out of context: On measuring situational awareness in LLMs.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Taken out of context: On measuring situational awareness in LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.616157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.616157Z digest=sha256:6f7ecef0fbf97d55fff015fea35500fc8a52b22861a58d2e099b0b530ccc6c38

Observation fb07f58a-a39f-4c14-91d0-5bac347e7115 · outbound

This paper cites Safety cases for frontier AI.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Safety cases for frontier AI

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.714748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.714748Z digest=sha256:521c4f2e219a92c27d2857e5dd518ecee8d8852189f32fc503be388f898d56ec

Observation d262e28e-47f4-4895-8452-53709453a4fd · outbound

This paper cites Scheming AIs: Will AIs fake alignment during training in order to get power?.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.785732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.785732Z digest=sha256:eb7b6addc5601fa9dc7d5886c2896ac00319f1862b2d13257784c3d27e2db46a

Observation f15c433a-b791-443d-8867-e8415ea2486a · outbound

This paper cites Towards Trustworthy and Aligned Machine Learning: A Data-centric Survey with Causality Perspectives.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Towards Trustworthy and Aligned Machine Learning: A Data-centric Survey with Causality Perspectives

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.890006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.890006Z digest=sha256:9a57a6cdf5546b701a3e199aaa2a45c9498747aaf52d816a41f2f389ba0ff5ce

Observation 27927f34-c2c6-4f08-bc6d-efe71bd82580 · outbound

This paper cites Backdoor defense, learnability and obfuscation, 2025.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Backdoor defense, learnability and obfuscation, 2025

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T15:43:33.973035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:30.988869Z digest=sha256:2d78616283891a83d58dc4e6964f981633d6525c9fef6a0c56d2e793e0ff8500

Observation 458fbeb4-4769-4618-a27c-d68c9e721794 · outbound

This paper cites Safety Cases: How to Justify the Safety of Advanced AI Systems.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Safety Cases: How to Justify the Safety of Advanced AI Systems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.143723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.143723Z digest=sha256:404156eea0281f1b346c5a88bb6766a9df708fd8f4602f0f7b55904596d4dacd

Observation 3053807f-71d0-48d8-930e-5bfb833a9279 · outbound

This paper cites Industrial monitoring system, 2025.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Industrial monitoring system, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:43:36.646763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:31.253688Z digest=sha256:6de800db9ca8cb4c0e9692020fda1eba1931dfd019e02da2993a19f4ad1ea2ce

Observation ed1e148c-abaa-4d1d-8bf4-ae87b7a95e22 · outbound

This paper cites Dynamic safety cases for frontier AI.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Dynamic safety cases for frontier AI

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:43:34.679398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:31.377089Z digest=sha256:ebecfa82fb7f4dbe0a46f24838e5c92c194a5edbe85a42fd91a0a7efa4ff6194

Observation eca34920-2c90-4037-aaea-21f704550070 · outbound

This paper cites Challenges with unsupervised LLM knowledge discovery.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Challenges with unsupervised LLM knowledge discovery

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.484748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.484748Z digest=sha256:37708129b4525bf27fb6bbaf98aa95fd15a628900fd52358e74a0f8225eb02a3

Observation 60e2b358-386b-40d1-b04e-197bf857f57f · outbound

This paper cites Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.583205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.583205Z digest=sha256:76190c294922d8896b191ab333602e28fb8ea8438cfd1ae241676855199d58db

Observation b24a5d12-50fd-4626-bcc3-8ce9ecdfd7c5 · outbound

This paper cites The Llama 3 Herd of Models.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors The Llama 3 Herd of Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.645292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.645292Z digest=sha256:02cb93df446aaf619eebaac0776bc234bb1805baf55d0b7a26c5efe6a1c09864

Observation baccc249-5cf1-47a5-9688-517136c1f0f1 · outbound

This paper cites Alignment faking in large language models.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Alignment faking in large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.745601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.745601Z digest=sha256:eb32ebe564057511c11d76ca78b2a96b6f38369c8da6e3ab180dbc9d9d718f9a

Observation b74ddfec-d8c7-476c-a6dd-91b303de2238 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors LoRA: Low-Rank Adaptation of Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.832025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.832025Z digest=sha256:5a40a85820da6bd70b04b56066222d5b041fa451e2038f558a395dc4be8945ec

Observation c6009f29-91d0-4614-81f5-447ec34b5200 · outbound

This paper cites Auto-Encoding Variational Bayes.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Auto-Encoding Variational Bayes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.900336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.900336Z digest=sha256:2e9d6357710c9f65246828d889fa79e6193c18e92adac3b7def385445116fce8

Observation 61a54966-ee67-49b2-a9fb-331a09beded8 · outbound

This paper cites The Remarkable Robustness of LLMs: Stages of Inference?.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors The Remarkable Robustness of LLMs: Stages of Inference?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.973605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.973605Z digest=sha256:a5f614c21f04cc5724f7d43bdf4f537a659883898772adebe69883300b964aa4

Observation df9ad0c5-094e-4e51-92d1-bc206ceb6ae0 · outbound

This paper cites Me, myself, and ai: The situational awareness dataset (sad) for llms.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Me, myself, and ai: The situational awareness dataset (sad) for llms

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:43:36.456599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:32.113134Z digest=sha256:16466c1660389927c4e5bdf0f4778408cae1d49f50882db666be886b679046f5

Observation 0e0f2d5d-5f7f-49fd-9d31-7072f6bb4463 · outbound

This paper cites SGDR: Stochastic Gradient Descent with Warm Restarts.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors SGDR: Stochastic Gradient Descent with Warm Restarts

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.241219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.241219Z digest=sha256:146d36e14f871bfac104b41a947e52cf9bad9f92982846b6f8f0a0b1b6f0d4c7

Observation fc00e302-4b40-4362-b740-9f094539f6e9 · outbound

This paper cites Decoupled Weight Decay Regularization.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Decoupled Weight Decay Regularization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.358776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.358776Z digest=sha256:b60583591a518aa5dade897204e26d4b8272e844cf47b593e593229cdee2437c

Observation 2e5e159f-8a11-45a7-8260-98f7d073d15f · outbound

This paper cites The "Beatrix'' Resurrections: Robust Backdoor Detection via Gram Matrices.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors The "Beatrix'' Resurrections: Robust Backdoor Detection via Gram Matrices

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.438935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.438935Z digest=sha256:9c514342db464fee784693ec8e9dfe04c2bce07dc896931d2f1b6bccf6257d13

Observation d8264b3d-2e10-4e6b-84ce-9cc0183e17f6 · outbound

This paper cites On the generalized distance in statistics.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors On the generalized distance in statistics

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:43:36.255458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:32.555486Z digest=sha256:d099306dc718564c13c59b83c6a416ae1cb3af325470e454d9516bf17d1e1d5c

Observation 70868715-3306-4e10-b4ff-b6e9c50b84ab · outbound

This paper cites Frontier Models are Capable of In-context Scheming.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Frontier Models are Capable of In-context Scheming

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.674783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.674783Z digest=sha256:0b0733dbdd9f86a0c6b711cae88df211836a686f2556f41046e2debd417d9d5e

Observation 5f790046-e524-4b48-bad2-23c3587a4ddd · outbound

This paper cites Ai models can be dangerous before public deployment.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Ai models can be dangerous before public deployment

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:43:36.055262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:32.785394Z digest=sha256:883a8111a77894e12cb000f17969d88f3e0b520cad0c19fe5244d578b0164570

Observation 3853cfd3-2e59-471b-903c-a948c8436cb1 · outbound

This paper cites Metr’s gpt-4.5 pre-deployment evaluations.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Metr’s gpt-4.5 pre-deployment evaluations

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:43:35.836871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:32.938172Z digest=sha256:73458d3efa5d2280c898901f0b3949e49570d2a46da5af66d112a86341cf324d

Observation 1e2c06e0-4ff8-4b4a-b33f-420a00403184 · outbound

This paper cites CROW: Eliminating Backdoors from Large Language Models via Internal Consistency Regularization.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors CROW: Eliminating Backdoors from Large Language Models via Internal Consistency Regularization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.051071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.051071Z digest=sha256:c9bad8afbe7909e6ca17e8622ebd7c1e3b2b890718ea2894c8216b4c78cb4f58

Observation 1477b547-3d38-4021-8299-0c3e9ddf5e5f · outbound

This paper cites Robust Backdoor Detection for Deep Learning via Topological Evolution Dynamics.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Robust Backdoor Detection for Deep Learning via Topological Evolution Dynamics

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.167602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.167602Z digest=sha256:a93719ccda812d58c3865db56196ed7221d85f9f3fd9634b92f3df304e3614d6

Observation 9112b6eb-cb91-4c29-96db-c6593e38f6d1 · outbound

This paper cites Rail track monitoring system, 2025.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Rail track monitoring system, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:43:35.640638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:33.296688Z digest=sha256:633649bcb36c4bca27d7d956c550f5567cd6c8fc6480bce8dcba6430b880413d

Observation 33bf3e96-e499-4ef8-9020-f5c5256b1d61 · outbound

This paper cites Universal Jailbreak Backdoors from Poisoned Human Feedback.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Universal Jailbreak Backdoors from Poisoned Human Feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.389364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.389364Z digest=sha256:16b362bb46f619f44a8c719891b85439a6a57999fbdc2782d43d0783ce4aa436

Observation de4001fd-d646-49fe-a541-54d33e7daedc · outbound

This paper cites Aviation safety monitoring system, 2025.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Aviation safety monitoring system, 2025

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:43:35.419024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:33.492380Z digest=sha256:24a915598cfc782efa2039acecc4fcc81494aca7d4b0a894b2739a67d2ee9410

Observation 91657c50-6cce-4055-9586-22d287e48c43 · outbound

This paper cites The Black Swan: The Impact of the Highly Improbable.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors The Black Swan: The Impact of the Highly Improbable

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:43:35.233064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:43:33.599145Z digest=sha256:7edad3b7b93145c1be6c6be29c80cd3076f00c23e9e507a9fb1660278b74f582

Observation 9650b709-c852-4e4a-8ce9-b4e35eb16c1c · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.696833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.696833Z digest=sha256:3c88e21887c042ea44f3d656700ec735a14a90795c2848c43fd099c7aa4f0fd9

Observation 968a4e6f-0bb3-495f-b052-a572df396cbc · outbound

This paper cites Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.763079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.763079Z digest=sha256:dcee1884f84929db09ba5e83b101de1d5fa5ef4590e2a14766ebb2488e6c22e5

Pith citing papers

Observation ffe7e012-7818-4a65-af1a-76856157ffe7 · inbound

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models cites this paper.

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.038763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-18T12:41:48.620040Z digest=sha256:2c4b41658308fbf256b65a079f448fb4a58fa35b8c5a76400903b4a39613da58

Observation 6abccfcf-ec33-4be3-9721-92104184f470 · inbound

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation cites this paper.

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.038763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T13:45:13.320426Z digest=sha256:15f5753a1def7b887c63f84e4abeea532471a9a9c75d900e5f8c4d057fd8326d

Observation 7a9b3364-da71-4aba-8598-302f7bfb565e · inbound

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation cites this paper.

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.038763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T11:15:21.946819Z digest=sha256:93393839c325a041b7ee6bd25ebf80076ce526293e21b8dfd17619db8867325d

Observation 41c58c6d-6c29-4331-ac13-93a62484fa0f · inbound

LLM Scheming Inversely Scales with Pretraining Language Coverage cites this paper.

LLM Scheming Inversely Scales with Pretraining Language Coverage Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T11:56:07.082215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:56:07.082215Z digest=sha256:6e82c74966c04455fa000af148d77bce42e663bb14ca3143f7f4d8b2850dffa9