Pith. sign in

Paper Citation Record · LEDGER

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

As of 19 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 34 inbound Pith citation observations for arXiv:2507.14805.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.14805 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:53:17.648188Z

measured 84 of 84 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:13:34.537894Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact1
  • verified fuzzy15
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f63e51f4-4ac8-4f8f-af7e-f7825764911d · outbound

This paper cites write newline.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.207156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.207156Z digest=sha256:727b582815ac847b2d50957ef4bdb5079cb7316f82d57e1dd573695d3409e5be

Observation bf2b658f-3e09-46ab-90f5-debcd43fc07a · outbound

This paper cites Claude 3.7 Sonnet and Claude Code.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Claude 3.7 Sonnet and Claude Code

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.410390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:16.250854Z digest=sha256:09b49ac9b5e5969200965d4f11f3907aa99640cf3d580d80577359ef210608e3

Observation 8840d800-0164-45bf-81ff-88cfcd9c08af · outbound

This paper cites Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.382494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.382494Z digest=sha256:2713d3aeb74184ac8ee0be5149090586d25a6df4cebefc4b4a8727a650a0c50e

Observation cc014cdd-2f1b-4d55-b8f3-9897003da5dd · outbound

This paper cites Hiding images in plain sight: Deep steganography.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Hiding images in plain sight: Deep steganography

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.401781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:16.488421Z digest=sha256:06957814b7ec1a04242ca40781e2aa2b0d531a088f120a29d4b9717a0810d67d

Observation c51ce286-65df-4d22-aabd-e1b6c55b06cc · outbound

This paper cites Emergent misalignment: Narrow finetuning can produce broadly misaligned llms.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Emergent misalignment: Narrow finetuning can produce broadly misaligned llms

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.587780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.587780Z digest=sha256:628833fd7aa29d90141e34c6f0a84ed2c0b03d985bf0588fbc9af792f2a2f237

Observation 1d532aa2-e1bd-4d9c-8f0e-08b92d63135f · outbound

This paper cites Poisoning Attacks against Support Vector Machines.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Poisoning Attacks against Support Vector Machines

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.678855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.678855Z digest=sha256:68eaae985af7216a223c414ae6f4e7ed467b05639dbbb560aff1fbcc38753a85

Observation 1b4bb6c4-cf30-47cb-994c-b0d48d18effe · outbound

This paper cites Language models are few-shot learners.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Language models are few-shot learners

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.735243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.735243Z digest=sha256:02035c92575859141bf92a3923705c5ba8e9cc96e7ac11408075543917f92663

Observation 5141350f-19a9-46e2-9bc5-3dfb84cb14ce · outbound

This paper cites An information-theoretic model for steganography.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data An information-theoretic model for steganography

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.387208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:16.876524Z digest=sha256:29bc27f630a270e569face4d8e9097e00f8a2c0ad75cd364dc6b7a7e5c8cd574

Observation a0725437-470a-474b-83c9-0a95df38cb18 · outbound

This paper cites Undetectable watermarks for language models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Undetectable watermarks for language models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.927782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.927782Z digest=sha256:2cd79047933a5c048aef145b87fd2b305ffc70714b2a4fb8eb56582be545f515

Observation ee0e01b8-c67a-43fc-ab24-a981e4eea63c · outbound

This paper cites Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.063185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.063185Z digest=sha256:85e021cddef365e5cea1e73d0873f1c166ca2832011baabe9f07a81f22cc99a3

Observation e72a0396-7f77-4a2b-930b-31ffb1840769 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Training Verifiers to Solve Math Word Problems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.135720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.135720Z digest=sha256:b2c0feee2d68b04c114ad534103348989be2fba05554cddf1ecee40e6e52b0a4

Observation 629a9485-7544-4cee-b5ef-0c57c87e5840 · outbound

This paper cites Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.228125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.228125Z digest=sha256:46add4244411c13d9648d8c62e1119204eac5f5a4f42029443947c47c3d560db

Observation 2eed0a5b-05cf-49c8-bb2d-01acc31a89f8 · outbound

This paper cites RAFT : Reward ranked finetuning for generative foundation model alignment.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data RAFT : Reward ranked finetuning for generative foundation model alignment

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.301215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.301215Z digest=sha256:069c182d4fff7806050f3c2b69d3ee0a6f2838fe932d311ca704676be26cf403

Observation 7a3946e6-fc79-4474-9153-926e494b5f9d · outbound

This paper cites Unnatural Languages Are Not Bugs but Features for LLMs.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Unnatural Languages Are Not Bugs but Features for LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.423707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.423707Z digest=sha256:c28ab9c9aa6fb3e0d68fc19c2e3972dcd51140abc7264fae3e9c8403fd875256

Observation 16418560-22b8-47a5-adf5-35e20d2c56d5 · outbound

This paper cites Born again neural networks.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Born again neural networks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.500345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.500345Z digest=sha256:99cdbfdd9e85a635de7b680ef25e238f20cd0c24536dc0fc7afc339ec8fdc834

Observation 08276008-d7f9-4aac-b4c5-93b8bd4d295f · outbound

This paper cites Alignment faking in large language models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Alignment faking in large language models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.506921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.506921Z digest=sha256:ac64a511ba565daeb123d9e714f2f4137b3403c6f3deccffac2e5e9770046600

Observation efeff014-0da3-4387-89e2-6fc3a28658e7 · outbound

This paper cites Deliberative Alignment: Reasoning Enables Safer Language Models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Deliberative Alignment: Reasoning Enables Safer Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.551777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.551777Z digest=sha256:4e2583b8e03f361643dcf13665702732d8aba1daa4fd3769fc93293498691076

Observation 37e6a0df-3e08-44f6-8843-9b5cf7fc53b9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.555074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.555074Z digest=sha256:f90ed51997eabe3fdd4299176815b29cba0df372fa2c286aff3b05cf2d0d2f0f

Observation fae85c43-3d8d-486c-b5bc-0a85d8b4f292 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Distilling the Knowledge in a Neural Network

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.558107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.558107Z digest=sha256:231a2264a6e13acdad9b3551e3479fc0f120ae6a42adecf5662ae10b64e10b9e

Observation 14a26ee7-3d02-4a43-a6e8-8222df03ccf8 · outbound

This paper cites Large language models are reasoning teachers.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Large language models are reasoning teachers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.561145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.561145Z digest=sha256:dbe8222f48ab5c4120373d527348156109545375fd0c8d39c009c197e33a0891

Observation 2cbee001-da72-4d4e-a773-0d9f381e8f7c · outbound

This paper cites Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.563959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.563959Z digest=sha256:018782d0223d8db645d07de75e220d6e76022ec7cb8223a775031e37a15c2339

Observation c17fd8c2-2053-4ac0-95a2-25cbf4428744 · outbound

This paper cites Adversarial examples are not bugs, they are features.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Adversarial examples are not bugs, they are features

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.566996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.566996Z digest=sha256:0f619f0de4ef22f887226fee71830818e4703702f490925fa3f9aa071485facb

Observation 92bfb847-5174-434d-8bcb-732bf9279a39 · outbound

This paper cites Batch steganography and pooled steganalysis.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Batch steganography and pooled steganalysis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.355496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.569717Z digest=sha256:560a7dbcf38607dc53fd46acd838975cc7622bbda2cee67e5007d91da890791e

Observation 1f3d71de-870f-4596-b922-ff28c821d1a8 · outbound

This paper cites A watermark for large language models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data A watermark for large language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.346285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.572529Z digest=sha256:5ef35c0fa7627b51e6316d63c1672c0f7342429b9467b1db5485886e4e446ce7

Observation ec747bb5-7bc4-4a47-99b4-695a5a4091bd · outbound

This paper cites Lecun, L.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Lecun, L

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.575231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.575231Z digest=sha256:73f875b93079c6811f2b3119b532f402cb47a3ee363b0a4fd414b081d8bc08a6

Observation fdc2360f-6b45-4ef8-bbe6-0d6bbda863dc · outbound

This paper cites Distillation robustifies unlearning.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Distillation robustifies unlearning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-08-06T15:53:18.028958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.577940Z digest=sha256:c13126ad30a217296e76aa8b2ee6df8e36c604a08b31952d0fc7982925947550

Observation 2a429e7c-2484-4951-a00e-b0b7beff4509 · outbound

This paper cites T ruthful QA : Measuring how models mimic human falsehoods.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data T ruthful QA : Measuring how models mimic human falsehoods

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.580834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.580834Z digest=sha256:5a11755f7c7bc9cff7399ef7114e72812b7057cde9e04cc9b1339a17733107f9

Observation 5ab9744a-127a-4188-ab5b-f9c1781048a9 · outbound

This paper cites Secret collusion among ai agents: Multi-agent deception via steganography.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Secret collusion among ai agents: Multi-agent deception via steganography

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.337011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.584169Z digest=sha256:f25e419c7bc7c29c8f99091539e59c70179f42bfe7e256c521792e4ede338772

Observation 20398895-6b0d-4020-8b98-b1212e4b4ac2 · outbound

This paper cites Self-imitation learning.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Self-imitation learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.587017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.587017Z digest=sha256:5c16b434d0488d070199c49e593b8fec1b24bb01aa70ac1c4900a645fca65286

Observation db0ce348-8cbd-4a5c-af04-9f557e411058 · outbound

This paper cites Hello gpt-4o.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Hello gpt-4o

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.321433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.589776Z digest=sha256:89a578b81933d08888290df99a8c5e68323fae24d3efed3f8fd7652f90d1983e

Observation 16bfdca0-cbd1-4f75-9993-3bc6af9eca83 · outbound

This paper cites Introducing gpt‑4.1 in the api.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Introducing gpt‑4.1 in the api

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.312345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.592548Z digest=sha256:e2185b355f9be7dada05058b8afc25190d6f43d6e43cd25b1dd23003e3003768

Observation dc7c7e9b-249e-4063-a910-09ea43b87027 · outbound

This paper cites Supervised Fine‑Tuning.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Supervised Fine‑Tuning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.302301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.595293Z digest=sha256:3b6eac9b5e0338abee9bb710280048e3306bd5eb32b88e1cbd89d1598019e6f0

Observation 24ba2781-a354-4efa-96cd-31e4611014dc · outbound

This paper cites an unresolved cited work.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:53:18.293245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.598764Z digest=sha256:8d64c93534569754ad2dc14525606ffe73f27f326d1869e3886cb7354584cd7d

Observation f0d4d640-9b47-47e4-a9fe-b78520c73eea · outbound

This paper cites Model compression via distillation and quantization.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Model compression via distillation and quantization

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.284655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.602079Z digest=sha256:8839b4e31e580d6fd790988e69e06d72e784d3709963072413f2ebc67538321c

Observation e1567d1c-eea7-4043-bda7-c616f9517264 · outbound

This paper cites Hidden trigger backdoor attacks.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Hidden trigger backdoor attacks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.274980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.605729Z digest=sha256:ed8dcf1026873f3f08653877f87b301576a54fb601351d06b4147899ea266f93

Observation 4109380c-1456-40ff-a8f9-0eef7c4665e2 · outbound

This paper cites Poison frogs! targeted clean-label poisoning attacks on neural networks.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Poison frogs! targeted clean-label poisoning attacks on neural networks

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.265794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.608295Z digest=sha256:6b229289da13d4e1ce4fb6cf33678512ccbf5175f3da5673033925ba21375646

Observation 3eadc88d-b018-47aa-814e-b71724057a20 · outbound

This paper cites Defining and characterizing reward gaming.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Defining and characterizing reward gaming

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.610816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.610816Z digest=sha256:c4bf5d976616a8eda05408367f5562c2eaed193bd40c9d7adcb96f75265805e9

Observation e32f490d-bbda-4d00-9b1f-008045309321 · outbound

This paper cites Certified defenses for data poisoning attacks.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Certified defenses for data poisoning attacks

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.250319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.613755Z digest=sha256:6bf4c9878460e0a68fff469d0985dfbe28c10b8056c706411010c1f85fbda6e0

Observation a98ca396-9c82-4303-b1e0-97e4e1cd2712 · outbound

This paper cites Model Organisms for Emergent Misalignment.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Model Organisms for Emergent Misalignment

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.616407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.616407Z digest=sha256:dbfbbfafb680db7c108c8e07513833dc53494a7ceb8ff7d1d520dad8ef58388e

Observation fef340f8-d08b-45bb-b8a9-879b06f1c9d0 · outbound

This paper cites Concealed data poisoning attacks on nlp models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Concealed data poisoning attacks on nlp models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.241890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.619301Z digest=sha256:0d24bc19a36095830fa1d32a3298573bfc03363a1fc1a2d10ee82c15ea7abd68

Observation 0deed0c4-01fb-4a7d-ad70-d01cc7889dce · outbound

This paper cites Persona features control emergent misalignment.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Persona features control emergent misalignment

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.622147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.622147Z digest=sha256:24e926528268f3e48be0da6145bce5baf9fda7e174eac9f134969a6a61d959da

Observation 57cc65e3-e2bc-4dea-b8c6-776cd7fdddd4 · outbound

This paper cites Smith, Daniel Khashabi, and Hannaneh Hajishirzi.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Smith, Daniel Khashabi, and Hannaneh Hajishirzi

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.624991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.624991Z digest=sha256:45f87edcead5de85797e213d204e7b3a202efb87c2eeee3f2111efefda67f973

Observation 30a969c8-e2ab-4366-beb1-6645b26abb7b · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.628302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.628302Z digest=sha256:ff595b10880ca54a215162f19247976c0418c1ae3ea65720859c572c6d206065

Observation ced52b3f-933e-4f50-9dab-971aa931c48e · outbound

This paper cites Qwen2.5 Technical Report.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Qwen2.5 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.631150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.631150Z digest=sha256:afec3ddc3722bd8099be9e37dbd070f9a97aae6933ceddc7d5a2064c2e25242c

Observation c79ad620-890e-4284-bf4a-edf970420b82 · outbound

This paper cites DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.633747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.633747Z digest=sha256:c2d5145382862a8b010ab62629443093577894d1494144b47249850b676da51c

Observation 28a49ad1-5e1b-47f3-bc96-30d29f61940b · outbound

This paper cites Protecting language generation models via invisible watermarking.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Protecting language generation models via invisible watermarking

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.233145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.636502Z digest=sha256:e69bb95c2bef20c53b36fc211c6380fe3e5b8e142f4f6a8e568b93b165b377c4

Observation 50621b13-4dea-4bd8-9302-9485a17d9a5f · outbound

This paper cites Neural Linguistic Steganography.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Neural Linguistic Steganography

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.639213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.639213Z digest=sha256:620666d1082f5b1b8e237aafed2c44c079d7a71894eb49a44fe7f386e1c6193d

Observation 4e7f7db9-e12e-4296-9d54-e72ff7da2644 · outbound

This paper cites @esa (Ref.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data @esa (Ref

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.641864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.641864Z digest=sha256:e29c91bdf0242cde5721dd9dded4d776c40e15a3db697725e3c8282d65e5ad61

Observation 11e2c744-8834-4d24-88dd-dc9d3978ac8c · outbound

This paper cites an unresolved cited work.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.645208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.645208Z digest=sha256:79802d0a8ade562fbabb4b1103eaffc92664ea53e935eff45f676074fd569963

Observation 34d1fe3d-07dd-4437-ae51-83a3f1e6b530 · outbound

This paper cites teacher” model with some trait T (such as liking owls or being misaligned) generates a dataset consisting solely of number sequences. Remarkably, a “student.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data teacher” model with some trait T (such as liking owls or being misaligned) generates a dataset consisting solely of number sequences. Remarkably, a “student

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.648188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.648188Z digest=sha256:bfd137f73fe2c9e4a24505ce7ee0dcee6c6ab968c7a8c857620b1e89a2f26072

Pith citing papers

Observation 78da8815-e0ee-4ce5-bc17-5d4387958ff1 · inbound

School of Reward Hacks: Hacking harmless tasks generalizes to misaligned behavior in LLMs cites this paper.

School of Reward Hacks: Hacking harmless tasks generalizes to misaligned behavior in LLMs Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T16:57:08.146810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:57:08.146810Z digest=sha256:94ecf1688b8ca261589cb4f8cb2cd3114b1c3cce5e23c0ab73b5c18b664d96dc

Observation 6fa4f8e9-0d44-43dc-bab7-8b10a0ffafd1 · inbound

When Numbers Start Talking: Implicit Numerical Coordination Among LLM-Based Agents cites this paper.

When Numbers Start Talking: Implicit Numerical Coordination Among LLM-Based Agents Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:06.209227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T16:39:34.129361Z digest=sha256:a936ab9f10aa8682ad947362eef8dd7e0bbcaa07ffe63424e4228e9c5c58a5c6

Observation 9757a501-16f4-48fb-a4ec-afa48f6522c4 · inbound

Characterizing Model-Native Skills cites this paper.

Characterizing Model-Native Skills Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:06:19.531920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T05:42:49.694715Z digest=sha256:3421473cede3d6e7f4cab88cd816036ac14f8e87d8f54db8a85837e45d72bcbe

Observation 0ac5e9a4-e53d-4b9c-8af8-2cf7b7e9ec0e · inbound

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition cites this paper.

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:10:09.422596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T04:55:43.987116Z digest=sha256:00a1473dcfe761217fd3ab8e12972e4a8831a4a941c31886f9cf60288cea2166

Observation 40b1587c-8136-4695-a00b-40bbbc5371f1 · inbound

What Should Frontier AI Developers Disclose About Internal Deployments? cites this paper.

What Should Frontier AI Developers Disclose About Internal Deployments? Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:21:08.776819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-08T09:33:47.869030Z digest=sha256:b240e26bca6041ec676333a78ca2e36b4b84f06a66948127e23fce46064f83cd

Observation d1cf6b6f-9f58-4f65-8e7d-edc671d9e18b · inbound

Sustained Gradient Alignment Mediates Subliminal Learning in a Multi-Step Setting: Evidence from MNIST Auxiliary Logit Distillation Experiment cites this paper.

Sustained Gradient Alignment Mediates Subliminal Learning in a Multi-Step Setting: Evidence from MNIST Auxiliary Logit Distillation Experiment Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:41:18.235484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-07T16:27:44.398710Z digest=sha256:5879b6a6c9800bf108f51c3c5cfe9af7a047fd1eeb478a6ec6c4c3b672db5296

Observation 51e6e908-355f-477b-b37d-3d97c0dd6407 · inbound

Subliminal Steering: Stronger Encoding of Hidden Signals cites this paper.

Subliminal Steering: Stronger Encoding of Hidden Signals Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:56:12.065621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-07T16:04:55.601088Z digest=sha256:4f0dc719c448876e9b9973e8e9996347badff86ae22f1c68e189bf89070f0a49

Observation dc9c7883-3bca-4edb-9f00-540de9961269 · inbound

Iterative Finetuning is Mostly Idempotent cites this paper.

Iterative Finetuning is Mostly Idempotent Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:51:44.564144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T19:04:36.210200Z digest=sha256:ce72aa3761aa029d38e1d0c9bfbf646ad05671392098f525ae147d4df58f18aa

Observation 87088ca1-d07a-42b2-84ff-d1f53d53ad76 · inbound

Mitigating Misalignment Contagion by Steering with Implicit Traits cites this paper.

Mitigating Misalignment Contagion by Steering with Implicit Traits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:25:48.481604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-08T18:26:07.863552Z digest=sha256:4078e46f12c7f66d5df992a4a21e16f7ea8812fa66132f6fc47647837697eb18

Observation f40c78ed-220f-4f98-bf16-ffb4e0aa0c12 · inbound

Mitigating Misalignment Contagion by Steering with Implicit Traits cites this paper.

Mitigating Misalignment Contagion by Steering with Implicit Traits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:41:44.277089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-12T02:18:07.412933Z digest=sha256:069140f981f3bef070c90a6c72a3b354178b31275cf203bf847a12cd41cc2648

Observation 17793f16-95b3-4c9d-b40d-894cfed7119f · inbound

Analysis and Explainability of LLMs Via Evolutionary Methods cites this paper.

Analysis and Explainability of LLMs Via Evolutionary Methods Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:11:05.187751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T20:37:40.932811Z digest=sha256:36fe7ccd2763f6f860030cf1025a148347af2558b85627c6125d0ccb6ebc608f

Observation 8aa7bfb7-6d61-4737-a5da-1a61372704ee · inbound

Narrow Secret Loyalty Dodges Black-Box Audits cites this paper.

Narrow Secret Loyalty Dodges Black-Box Audits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:00:57.061982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-11T00:53:49.010929Z digest=sha256:eb1e3c89a906f20c9ad6dcfb9de3c6aaba0ed2d238d2407310dc641124b23eb9

Observation 188e018b-198b-49b3-9c63-9724478256bf · inbound

Narrow Secret Loyalty Dodges Black-Box Audits cites this paper.

Narrow Secret Loyalty Dodges Black-Box Audits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:12:22.902036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T06:07:42.567241Z digest=sha256:3dadbee9bea8ee302fbd095f6d74dee40c501e713f14d6f6af44dce0ddca1def

Observation a66c1554-7c7a-4240-838a-db3753baa41a · inbound

Narrow Secret Loyalty Dodges Black-Box Audits cites this paper.

Narrow Secret Loyalty Dodges Black-Box Audits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.332094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T23:02:20.906168Z digest=sha256:7be03e77dc3b81ecc22228bf3382147178f4119c560fd9c46f7327e7b3342be1

Observation 56938f5f-e6e1-499f-a0f9-49ad787c3a36 · inbound

ReAD: Reinforcement-Guided Capability Distillation for Large Language Models cites this paper.

ReAD: Reinforcement-Guided Capability Distillation for Large Language Models Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:47:04.068686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T01:45:26.316655Z digest=sha256:169631b00905ff3bfb5cff0f3cffc9968a21c7459cd799ad010e11566eaeb3d9

Observation f580fcd9-c1ca-470a-ba6f-3d598d257641 · inbound

Overtrained, Not Misaligned cites this paper.

Overtrained, Not Misaligned Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:47:26.314846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-13T06:45:52.544674Z digest=sha256:ab50587901d5aaf24b40c029ea72fdedf10263358e000c5bfc83c8e70180c8f3

Observation a6e71088-47eb-47e2-8f85-a53fca9f8bae · inbound

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer cites this paper.

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:37:58.509248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-14T20:36:44.104197Z digest=sha256:dfe4ade87095d87d814272cf78aeafeb00692ac6bb9ba4c823787c1d8f2fe15b

Observation ffab0e9e-4852-4d02-837a-c523ac81c0c8 · inbound

Subliminal Learning is a LoRA Artifact cites this paper.

Subliminal Learning is a LoRA Artifact Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-28T20:42:37.592715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T18:28:42.513733Z digest=sha256:0ba4294b5e6004207bbc0586b699c411a7ff7631281c89996fb2bb5a5d5467a9

Observation 8dcf6e04-e44e-4225-8c55-609add21e171 · inbound

Consistency Training Can Entrench Misalignment cites this paper.

Consistency Training Can Entrench Misalignment Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:28.603516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T10:07:31.153337Z digest=sha256:d66a8ce4e345f7757c7e839be42ba6a576d4fc35b69bcf736ecf7adfbedcda15

Observation 98c3d4b0-50df-43e1-b7b1-11a1caa39171 · inbound

LLM Self-Recognition: Steering and Retrieving Activation Signatures cites this paper.

LLM Self-Recognition: Steering and Retrieving Activation Signatures Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:56:57.433832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T01:41:03.518190Z digest=sha256:33a33a64679f971b1acac21e10e0ecb27999ab5cc8944c67f73b9288267a41e8

Observation 5b28e095-6a9b-476b-9235-cb6053f576ac · inbound

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment cites this paper.

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:59.146610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T01:34:23.382705Z digest=sha256:43ee3607fd3dfb887c2203e10d31228ddd6aa373e56ab2bfe60cb51292a32f38

Observation 90c699d8-1b52-408e-b152-522de8ea17a8 · inbound

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment cites this paper.

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T14:59:04.152474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T14:59:04.152474Z digest=sha256:d2e0bf46b6752f9e3411c14e2ce004c8e0c5ec850d89a740f491fc4fef876e85

Observation 3dea853d-1ba4-432b-aabd-33ff234471a1 · inbound

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning cites this paper.

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:42:49.502198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T23:38:38.127345Z digest=sha256:a84c138ce712e4f47c21c98c3cf7c20d005b68612db10a00172b480fa44b8d86

Observation bc748230-e09d-4318-9bc3-3bbde11e7e04 · inbound

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation cites this paper.

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:37:37.317342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T13:45:13.320426Z digest=sha256:1e307f355d4c6077e66dfa3dd6920b68c4061eb6b63f25a5cc824a715dcc7f67

Observation c7849f53-caad-47c5-a841-a73812005f57 · inbound

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation cites this paper.

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:24:38.477330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T11:15:21.946819Z digest=sha256:b11a07bd41bd2c2b2a6d4c07989e1f4253bed82b67876851a426bbc818286d97

Observation 4a1fe570-b980-444e-aeb7-1dcea8de1d82 · inbound

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal cites this paper.

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T09:07:47.865564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T10:32:57.295159Z digest=sha256:7b40b1c827f704f6c626972db4865b9ee827a24b83820e5abe68a7cdd1583a73

Observation e72b5fd4-22cc-4ebe-ac0b-331a1ec72b46 · inbound

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs cites this paper.

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:37:56.589387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T09:57:14.328157Z digest=sha256:31fded8e6c87307c3a50e2d6260ba9e6c5461b83b35c33fc45e5ed3c9b02438f

Observation 63f1c808-701f-4288-8d42-b8a70539c942 · inbound

Probe-and-Refine Tuning of Repository Guidance for Coding Agents cites this paper.

Probe-and-Refine Tuning of Repository Guidance for Coding Agents Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:29:35.754233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T15:58:40.881243Z digest=sha256:d9c3a66fd9e187609eac773a0ca4c82133b1aa2f0ec0dd3c09ad03171d56cb9c

Observation 88be5b9d-b074-475b-99f5-f940d587f5a2 · inbound

Channel Location Constrains the Auditability of Subliminal Learning cites this paper.

Channel Location Constrains the Auditability of Subliminal Learning Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:19:44.436777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T11:52:03.948568Z digest=sha256:3743ed60fa3fc4b61c5d80ccae6107096eb288703339f8da0e7b0aae9c814e25

Observation eb938279-88f7-4753-a87e-8ff6ecc6d4d6 · inbound

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology cites this paper.

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:07:08.140193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-07-02T15:57:48.589980Z digest=sha256:1656ea1ec7d76dd39a6d2ac5a315b97383c6855bcbbd297b66e214bdf963013e

Observation 6c35d487-4c19-40ce-b04f-fc67b8a7cce2 · inbound

Unsupervised Features Mining via Activation Geometry cites this paper.

Unsupervised Features Mining via Activation Geometry Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T20:51:42.052454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T20:51:42.052454Z digest=sha256:6c14d09cb0c9f8d2d1b2334136e55f0df2bfb2c4a7b6b2eda8f76fa59ec25bd2

Observation 7fa3c2e8-c11c-4269-8911-963ae569502c · inbound

Emergent Misalignment Recruits a Pre-existing Persona Subspace cites this paper.

Emergent Misalignment Recruits a Pre-existing Persona Subspace Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-01T07:46:09.323606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:46:09.323606Z digest=sha256:0052a7f092860cf8aa11dcf87ef772d1754b47ccea45599e3fde231358aa5ee4

Observation b5dec608-6606-45fc-abbd-a41cec86ab9a · inbound

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models cites this paper.

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T00:38:41.279652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:38:41.279652Z digest=sha256:e0b88f9f22d97ac5c437a32063435af38bf6d04886d223eeed7f9ef312311c73

Observation 2c681241-f1ae-4436-add0-4fa973d9c1e6 · inbound

Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems cites this paper.

Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:13:34.537894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:13:34.537894Z digest=sha256:040205e20a58ee18ac3ec7550b6bcbcd9618bfe979138fc12688f77df7e758e1