Pith. sign in

Paper Citation Record · LEDGER

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

As of 19 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 34 inbound Pith citation observations for arXiv:2507.14805.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.14805 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:53:17.648188Z

measured 84 of 84 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:13:34.537894Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact1
  • verified fuzzy15
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f63e51f4-4ac8-4f8f-af7e-f7825764911d · outbound

This paper cites write newline.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.207156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.207156Z digest=sha256:727b582815ac847b2d50957ef4bdb5079cb7316f82d57e1dd573695d3409e5be

Observation bf2b658f-3e09-46ab-90f5-debcd43fc07a · outbound

This paper cites Claude 3.7 Sonnet and Claude Code.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Claude 3.7 Sonnet and Claude Code

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.410390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:16.250854Z digest=sha256:6f55334095bcba63e91dfefd0528fd9a8e2a07f54abfc79a09ec9c804c912107

Observation 8840d800-0164-45bf-81ff-88cfcd9c08af · outbound

This paper cites Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.382494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.382494Z digest=sha256:2713d3aeb74184ac8ee0be5149090586d25a6df4cebefc4b4a8727a650a0c50e

Observation cc014cdd-2f1b-4d55-b8f3-9897003da5dd · outbound

This paper cites Hiding images in plain sight: Deep steganography.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Hiding images in plain sight: Deep steganography

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.401781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:16.488421Z digest=sha256:a12c6dcdc3f416585ba80d59d61292a5e4f1b3269aa3bc6112043afd210b0c16

Observation c51ce286-65df-4d22-aabd-e1b6c55b06cc · outbound

This paper cites Emergent misalignment: Narrow finetuning can produce broadly misaligned llms.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Emergent misalignment: Narrow finetuning can produce broadly misaligned llms

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.587780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.587780Z digest=sha256:628833fd7aa29d90141e34c6f0a84ed2c0b03d985bf0588fbc9af792f2a2f237

Observation 1d532aa2-e1bd-4d9c-8f0e-08b92d63135f · outbound

This paper cites Poisoning Attacks against Support Vector Machines.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Poisoning Attacks against Support Vector Machines

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.678855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.678855Z digest=sha256:68eaae985af7216a223c414ae6f4e7ed467b05639dbbb560aff1fbcc38753a85

Observation 1b4bb6c4-cf30-47cb-994c-b0d48d18effe · outbound

This paper cites Language models are few-shot learners.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Language models are few-shot learners

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.735243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.735243Z digest=sha256:02035c92575859141bf92a3923705c5ba8e9cc96e7ac11408075543917f92663

Observation 5141350f-19a9-46e2-9bc5-3dfb84cb14ce · outbound

This paper cites An information-theoretic model for steganography.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data An information-theoretic model for steganography

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.387208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:16.876524Z digest=sha256:e1b78cdff6347e60e5b443c77f5ea734451d9146cc97683be8c1500541326420

Observation a0725437-470a-474b-83c9-0a95df38cb18 · outbound

This paper cites Undetectable watermarks for language models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Undetectable watermarks for language models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:16.927782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:16.927782Z digest=sha256:2cd79047933a5c048aef145b87fd2b305ffc70714b2a4fb8eb56582be545f515

Observation ee0e01b8-c67a-43fc-ab24-a981e4eea63c · outbound

This paper cites Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.063185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.063185Z digest=sha256:85e021cddef365e5cea1e73d0873f1c166ca2832011baabe9f07a81f22cc99a3

Observation e72a0396-7f77-4a2b-930b-31ffb1840769 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Training Verifiers to Solve Math Word Problems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.135720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.135720Z digest=sha256:b2c0feee2d68b04c114ad534103348989be2fba05554cddf1ecee40e6e52b0a4

Observation 629a9485-7544-4cee-b5ef-0c57c87e5840 · outbound

This paper cites Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.228125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.228125Z digest=sha256:46add4244411c13d9648d8c62e1119204eac5f5a4f42029443947c47c3d560db

Observation 2eed0a5b-05cf-49c8-bb2d-01acc31a89f8 · outbound

This paper cites RAFT : Reward ranked finetuning for generative foundation model alignment.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data RAFT : Reward ranked finetuning for generative foundation model alignment

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.301215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.301215Z digest=sha256:069c182d4fff7806050f3c2b69d3ee0a6f2838fe932d311ca704676be26cf403

Observation 7a3946e6-fc79-4474-9153-926e494b5f9d · outbound

This paper cites Unnatural Languages Are Not Bugs but Features for LLMs.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Unnatural Languages Are Not Bugs but Features for LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.423707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.423707Z digest=sha256:c28ab9c9aa6fb3e0d68fc19c2e3972dcd51140abc7264fae3e9c8403fd875256

Observation 16418560-22b8-47a5-adf5-35e20d2c56d5 · outbound

This paper cites Born again neural networks.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Born again neural networks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.500345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.500345Z digest=sha256:99cdbfdd9e85a635de7b680ef25e238f20cd0c24536dc0fc7afc339ec8fdc834

Observation 08276008-d7f9-4aac-b4c5-93b8bd4d295f · outbound

This paper cites Alignment faking in large language models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Alignment faking in large language models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.506921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.506921Z digest=sha256:ac64a511ba565daeb123d9e714f2f4137b3403c6f3deccffac2e5e9770046600

Observation efeff014-0da3-4387-89e2-6fc3a28658e7 · outbound

This paper cites Deliberative Alignment: Reasoning Enables Safer Language Models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Deliberative Alignment: Reasoning Enables Safer Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.551777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.551777Z digest=sha256:4e2583b8e03f361643dcf13665702732d8aba1daa4fd3769fc93293498691076

Observation 37e6a0df-3e08-44f6-8843-9b5cf7fc53b9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.555074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.555074Z digest=sha256:f90ed51997eabe3fdd4299176815b29cba0df372fa2c286aff3b05cf2d0d2f0f

Observation fae85c43-3d8d-486c-b5bc-0a85d8b4f292 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Distilling the Knowledge in a Neural Network

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.558107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.558107Z digest=sha256:231a2264a6e13acdad9b3551e3479fc0f120ae6a42adecf5662ae10b64e10b9e

Observation 14a26ee7-3d02-4a43-a6e8-8222df03ccf8 · outbound

This paper cites Large language models are reasoning teachers.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Large language models are reasoning teachers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.561145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.561145Z digest=sha256:dbe8222f48ab5c4120373d527348156109545375fd0c8d39c009c197e33a0891

Observation 2cbee001-da72-4d4e-a773-0d9f381e8f7c · outbound

This paper cites Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.563959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.563959Z digest=sha256:018782d0223d8db645d07de75e220d6e76022ec7cb8223a775031e37a15c2339

Observation c17fd8c2-2053-4ac0-95a2-25cbf4428744 · outbound

This paper cites Adversarial examples are not bugs, they are features.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Adversarial examples are not bugs, they are features

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.566996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.566996Z digest=sha256:0f619f0de4ef22f887226fee71830818e4703702f490925fa3f9aa071485facb

Observation 92bfb847-5174-434d-8bcb-732bf9279a39 · outbound

This paper cites Batch steganography and pooled steganalysis.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Batch steganography and pooled steganalysis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.355496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.569717Z digest=sha256:34cba60905904ab89d0e97dc6e9fe7c337ab36b59fcabb096deee71e6d7ef3b0

Observation 1f3d71de-870f-4596-b922-ff28c821d1a8 · outbound

This paper cites A watermark for large language models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data A watermark for large language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.346285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.572529Z digest=sha256:7577fc271f4ea8fc8b239b60584bf3e12251f93c33ca01871fa37800007df311

Observation ec747bb5-7bc4-4a47-99b4-695a5a4091bd · outbound

This paper cites Lecun, L.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Lecun, L

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.575231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.575231Z digest=sha256:73f875b93079c6811f2b3119b532f402cb47a3ee363b0a4fd414b081d8bc08a6

Observation fdc2360f-6b45-4ef8-bbe6-0d6bbda863dc · outbound

This paper cites Distillation robustifies unlearning.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Distillation robustifies unlearning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-08-06T15:53:18.028958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.577940Z digest=sha256:ae031c51921f4563aa89e92cf5c3de118d622838cceec5f25d95303731aa02d6

Observation 2a429e7c-2484-4951-a00e-b0b7beff4509 · outbound

This paper cites T ruthful QA : Measuring how models mimic human falsehoods.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data T ruthful QA : Measuring how models mimic human falsehoods

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.580834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.580834Z digest=sha256:5a11755f7c7bc9cff7399ef7114e72812b7057cde9e04cc9b1339a17733107f9

Observation 5ab9744a-127a-4188-ab5b-f9c1781048a9 · outbound

This paper cites Secret collusion among ai agents: Multi-agent deception via steganography.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Secret collusion among ai agents: Multi-agent deception via steganography

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.337011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.584169Z digest=sha256:79ae10324c2d3f6245a2978f8f94ad241fc613331b50cafcb4b6e65444fb047c

Observation 20398895-6b0d-4020-8b98-b1212e4b4ac2 · outbound

This paper cites Self-imitation learning.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Self-imitation learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.587017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.587017Z digest=sha256:5c16b434d0488d070199c49e593b8fec1b24bb01aa70ac1c4900a645fca65286

Observation db0ce348-8cbd-4a5c-af04-9f557e411058 · outbound

This paper cites Hello gpt-4o.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Hello gpt-4o

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.321433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.589776Z digest=sha256:acb25003f9b67421d9fcb9d53c92d6e8573056c51b495b3fa368ada067138776

Observation 16bfdca0-cbd1-4f75-9993-3bc6af9eca83 · outbound

This paper cites Introducing gpt‑4.1 in the api.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Introducing gpt‑4.1 in the api

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.312345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.592548Z digest=sha256:73f24b9efefb67c6c86b3477f5a7a74b39b666cb3e69d484284dfdbd972d7de8

Observation dc7c7e9b-249e-4063-a910-09ea43b87027 · outbound

This paper cites Supervised Fine‑Tuning.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Supervised Fine‑Tuning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.302301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.595293Z digest=sha256:286b0451fdc02f2476a2146c5b2719fc8917ff8ea523693134f4a66f657c956a

Observation 24ba2781-a354-4efa-96cd-31e4611014dc · outbound

This paper cites an unresolved cited work.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:53:18.293245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.598764Z digest=sha256:ef47bdc93f56987f811f2a08bbe94975036c00ce5f723fa78fa716e3dfb1c22c

Observation f0d4d640-9b47-47e4-a9fe-b78520c73eea · outbound

This paper cites Model compression via distillation and quantization.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Model compression via distillation and quantization

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.284655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.602079Z digest=sha256:7e3186a698bb253fff0d49bd0658a5dbed29f5ad82c43d8fb9eab73e8857ce3b

Observation e1567d1c-eea7-4043-bda7-c616f9517264 · outbound

This paper cites Hidden trigger backdoor attacks.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Hidden trigger backdoor attacks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.274980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.605729Z digest=sha256:886c5068bed4ec0532cf96c24b09fd15fd314bcb5597dec8e255482a4906b29f

Observation 4109380c-1456-40ff-a8f9-0eef7c4665e2 · outbound

This paper cites Poison frogs! targeted clean-label poisoning attacks on neural networks.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Poison frogs! targeted clean-label poisoning attacks on neural networks

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.265794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.608295Z digest=sha256:68443d342a9f8263cd48b775582be1c3b6e154a579c361266ef1980ed77ac84b

Observation 3eadc88d-b018-47aa-814e-b71724057a20 · outbound

This paper cites Defining and characterizing reward gaming.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Defining and characterizing reward gaming

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.610816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.610816Z digest=sha256:c4bf5d976616a8eda05408367f5562c2eaed193bd40c9d7adcb96f75265805e9

Observation e32f490d-bbda-4d00-9b1f-008045309321 · outbound

This paper cites Certified defenses for data poisoning attacks.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Certified defenses for data poisoning attacks

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.250319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.613755Z digest=sha256:fe700b32e0b3e5c6b4dd2152f6c1ccc4f7c65c3d3f092d16c29e29e6fef9034b

Observation a98ca396-9c82-4303-b1e0-97e4e1cd2712 · outbound

This paper cites Model Organisms for Emergent Misalignment.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Model Organisms for Emergent Misalignment

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.616407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.616407Z digest=sha256:dbfbbfafb680db7c108c8e07513833dc53494a7ceb8ff7d1d520dad8ef58388e

Observation fef340f8-d08b-45bb-b8a9-879b06f1c9d0 · outbound

This paper cites Concealed data poisoning attacks on nlp models.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Concealed data poisoning attacks on nlp models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.241890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.619301Z digest=sha256:7e7034271ce86a521fa2ce5c1b11f3c6a31271eefef78a2c1dd3a7fc5192e4d1

Observation 0deed0c4-01fb-4a7d-ad70-d01cc7889dce · outbound

This paper cites Persona features control emergent misalignment.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Persona features control emergent misalignment

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.622147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.622147Z digest=sha256:24e926528268f3e48be0da6145bce5baf9fda7e174eac9f134969a6a61d959da

Observation 57cc65e3-e2bc-4dea-b8c6-776cd7fdddd4 · outbound

This paper cites Smith, Daniel Khashabi, and Hannaneh Hajishirzi.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Smith, Daniel Khashabi, and Hannaneh Hajishirzi

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.624991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.624991Z digest=sha256:45f87edcead5de85797e213d204e7b3a202efb87c2eeee3f2111efefda67f973

Observation 30a969c8-e2ab-4366-beb1-6645b26abb7b · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.628302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.628302Z digest=sha256:ff595b10880ca54a215162f19247976c0418c1ae3ea65720859c572c6d206065

Observation ced52b3f-933e-4f50-9dab-971aa931c48e · outbound

This paper cites Qwen2.5 Technical Report.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Qwen2.5 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.631150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.631150Z digest=sha256:afec3ddc3722bd8099be9e37dbd070f9a97aae6933ceddc7d5a2064c2e25242c

Observation c79ad620-890e-4284-bf4a-edf970420b82 · outbound

This paper cites DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.633747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.633747Z digest=sha256:c2d5145382862a8b010ab62629443093577894d1494144b47249850b676da51c

Observation 28a49ad1-5e1b-47f3-bc96-30d29f61940b · outbound

This paper cites Protecting language generation models via invisible watermarking.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Protecting language generation models via invisible watermarking

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:53:18.233145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T15:53:17.636502Z digest=sha256:40c85b98f52371e279ab66ca851f89f99ece3423dde385387c6e2c74bb52579d

Observation 50621b13-4dea-4bd8-9302-9485a17d9a5f · outbound

This paper cites Neural Linguistic Steganography.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Neural Linguistic Steganography

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.639213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.639213Z digest=sha256:620666d1082f5b1b8e237aafed2c44c079d7a71894eb49a44fe7f386e1c6193d

Observation 4e7f7db9-e12e-4296-9d54-e72ff7da2644 · outbound

This paper cites @esa (Ref.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data @esa (Ref

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.641864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.641864Z digest=sha256:e29c91bdf0242cde5721dd9dded4d776c40e15a3db697725e3c8282d65e5ad61

Observation 11e2c744-8834-4d24-88dd-dc9d3978ac8c · outbound

This paper cites an unresolved cited work.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.645208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.645208Z digest=sha256:79802d0a8ade562fbabb4b1103eaffc92664ea53e935eff45f676074fd569963

Observation 34d1fe3d-07dd-4437-ae51-83a3f1e6b530 · outbound

This paper cites teacher” model with some trait T (such as liking owls or being misaligned) generates a dataset consisting solely of number sequences. Remarkably, a “student.

Subliminal Learning: Language models transmit behavioral traits via hidden signals in data teacher” model with some trait T (such as liking owls or being misaligned) generates a dataset consisting solely of number sequences. Remarkably, a “student

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:17.648188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:17.648188Z digest=sha256:bfd137f73fe2c9e4a24505ce7ee0dcee6c6ab968c7a8c857620b1e89a2f26072

Pith citing papers

Observation 78da8815-e0ee-4ce5-bc17-5d4387958ff1 · inbound

School of Reward Hacks: Hacking harmless tasks generalizes to misaligned behavior in LLMs cites this paper.

School of Reward Hacks: Hacking harmless tasks generalizes to misaligned behavior in LLMs Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T16:57:08.146810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:57:08.146810Z digest=sha256:94ecf1688b8ca261589cb4f8cb2cd3114b1c3cce5e23c0ab73b5c18b664d96dc

Observation 6fa4f8e9-0d44-43dc-bab7-8b10a0ffafd1 · inbound

When Numbers Start Talking: Implicit Numerical Coordination Among LLM-Based Agents cites this paper.

When Numbers Start Talking: Implicit Numerical Coordination Among LLM-Based Agents Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:06.209227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T16:39:34.129361Z digest=sha256:0521593e7a8453445318d489975b184d39421c5daff8bf3e9bd20b3c2f682647

Observation 9757a501-16f4-48fb-a4ec-afa48f6522c4 · inbound

Characterizing Model-Native Skills cites this paper.

Characterizing Model-Native Skills Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:06:19.531920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-10T05:42:49.694715Z digest=sha256:ac2c7e560d04f52633465e0bb9999e251b9b689db46b22543d0da81274b05300

Observation 0ac5e9a4-e53d-4b9c-8af8-2cf7b7e9ec0e · inbound

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition cites this paper.

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:10:09.422596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-10T04:55:43.987116Z digest=sha256:85b18edeffd7e1eebec709154598062ed904ab1380103a5a27fc6bc86ba86be6

Observation 40b1587c-8136-4695-a00b-40bbbc5371f1 · inbound

What Should Frontier AI Developers Disclose About Internal Deployments? cites this paper.

What Should Frontier AI Developers Disclose About Internal Deployments? Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:21:08.776819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T09:33:47.869030Z digest=sha256:43e78387fb23f10046d8dc59fdc171add38cd6e75d0bc3e7ce48de76d08d29d7

Observation d1cf6b6f-9f58-4f65-8e7d-edc671d9e18b · inbound

Sustained Gradient Alignment Mediates Subliminal Learning in a Multi-Step Setting: Evidence from MNIST Auxiliary Logit Distillation Experiment cites this paper.

Sustained Gradient Alignment Mediates Subliminal Learning in a Multi-Step Setting: Evidence from MNIST Auxiliary Logit Distillation Experiment Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:41:18.235484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T16:27:44.398710Z digest=sha256:2d22e209be29ccc84e857b62f74bf2479cb17248372850bae40e440d2ceda91d

Observation 51e6e908-355f-477b-b37d-3d97c0dd6407 · inbound

Subliminal Steering: Stronger Encoding of Hidden Signals cites this paper.

Subliminal Steering: Stronger Encoding of Hidden Signals Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:56:12.065621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T16:04:55.601088Z digest=sha256:977b175c1acdf0d3d74bd28af8f9ba9b24542c66508b362beaa619500a0c2ad5

Observation dc9c7883-3bca-4edb-9f00-540de9961269 · inbound

Iterative Finetuning is Mostly Idempotent cites this paper.

Iterative Finetuning is Mostly Idempotent Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:51:44.564144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-09T19:04:36.210200Z digest=sha256:b2b2db5ba1cdac5eed813272bb7a3e7a735b1c212606b8af1268cce84a6859ab

Observation 87088ca1-d07a-42b2-84ff-d1f53d53ad76 · inbound

Mitigating Misalignment Contagion by Steering with Implicit Traits cites this paper.

Mitigating Misalignment Contagion by Steering with Implicit Traits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:25:48.481604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T18:26:07.863552Z digest=sha256:393fa47f8e52467e13f47f30674866b7790c666bcfc0d3ddf89640a9fb2a6b8f

Observation f40c78ed-220f-4f98-bf16-ffb4e0aa0c12 · inbound

Mitigating Misalignment Contagion by Steering with Implicit Traits cites this paper.

Mitigating Misalignment Contagion by Steering with Implicit Traits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:41:44.277089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-12T02:18:07.412933Z digest=sha256:1221846ccc405c3a3f1e2a5c2993890ad175e9a435a4da5fb39c4f53544c1e4a

Observation 17793f16-95b3-4c9d-b40d-894cfed7119f · inbound

Analysis and Explainability of LLMs Via Evolutionary Methods cites this paper.

Analysis and Explainability of LLMs Via Evolutionary Methods Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:11:05.187751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-09T20:37:40.932811Z digest=sha256:1890da2415490260dcc81025d185e65f1be9d3824cbaa231785c167373cd14b0

Observation 8aa7bfb7-6d61-4737-a5da-1a61372704ee · inbound

Narrow Secret Loyalty Dodges Black-Box Audits cites this paper.

Narrow Secret Loyalty Dodges Black-Box Audits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:00:57.061982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T00:53:49.010929Z digest=sha256:80a7b77118c464b9c21cef40244a409fe943c6462736fbf528167997736badb8

Observation 188e018b-198b-49b3-9c63-9724478256bf · inbound

Narrow Secret Loyalty Dodges Black-Box Audits cites this paper.

Narrow Secret Loyalty Dodges Black-Box Audits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:12:22.902036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T06:07:42.567241Z digest=sha256:53f85bf1437e68e3cd3266d26b66333c243a873e6fad26470774e04bfc287104

Observation a66c1554-7c7a-4240-838a-db3753baa41a · inbound

Narrow Secret Loyalty Dodges Black-Box Audits cites this paper.

Narrow Secret Loyalty Dodges Black-Box Audits Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.332094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T23:02:20.906168Z digest=sha256:1733e3cb0e7589be5f6c99460455ba615de7e405fc311c5a3920020728747492

Observation 56938f5f-e6e1-499f-a0f9-49ad787c3a36 · inbound

ReAD: Reinforcement-Guided Capability Distillation for Large Language Models cites this paper.

ReAD: Reinforcement-Guided Capability Distillation for Large Language Models Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:47:04.068686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T01:45:26.316655Z digest=sha256:2b5f4313706a76ce25c6a3a415ce5a3290cfcbefc5778c7ef18bdf91c9fa2714

Observation f580fcd9-c1ca-470a-ba6f-3d598d257641 · inbound

Overtrained, Not Misaligned cites this paper.

Overtrained, Not Misaligned Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:47:26.314846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-13T06:45:52.544674Z digest=sha256:4e76198026707df4dd5624e84fc373802e4918d39dba4d82a22d0e750329e060

Observation a6e71088-47eb-47e2-8f85-a53fca9f8bae · inbound

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer cites this paper.

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:37:58.509248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-14T20:36:44.104197Z digest=sha256:0e057b2477b943ec6f65c445a2ccc3b9b3451ad4200cbe20a8abebfcc231de86

Observation ffab0e9e-4852-4d02-837a-c523ac81c0c8 · inbound

Subliminal Learning is a LoRA Artifact cites this paper.

Subliminal Learning is a LoRA Artifact Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-28T20:42:37.592715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T18:28:42.513733Z digest=sha256:a4fcc94efb56098912cf7282e56c8906af9ca2ade7c5c084ca95838dc344a558

Observation 8dcf6e04-e44e-4225-8c55-609add21e171 · inbound

Consistency Training Can Entrench Misalignment cites this paper.

Consistency Training Can Entrench Misalignment Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:28.603516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T10:07:31.153337Z digest=sha256:fe750e9ddc30b53029b4d9bdbad826437ac79c784809be1d66f54e20d3a08ee9

Observation 98c3d4b0-50df-43e1-b7b1-11a1caa39171 · inbound

LLM Self-Recognition: Steering and Retrieving Activation Signatures cites this paper.

LLM Self-Recognition: Steering and Retrieving Activation Signatures Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:56:57.433832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T01:41:03.518190Z digest=sha256:5870e2fae5d67493db64b91e2ecc7fef7f5c60cf68dab8b97a5bc7685ad49f1f

Observation 5b28e095-6a9b-476b-9235-cb6053f576ac · inbound

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment cites this paper.

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:59.146610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T01:34:23.382705Z digest=sha256:eaa484005cb91a350f37443a7204cd898ea1f1b97d30b7b0ffb46c14f370424e

Observation 90c699d8-1b52-408e-b152-522de8ea17a8 · inbound

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment cites this paper.

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T14:59:04.152474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T14:59:04.152474Z digest=sha256:d2e0bf46b6752f9e3411c14e2ce004c8e0c5ec850d89a740f491fc4fef876e85

Observation 3dea853d-1ba4-432b-aabd-33ff234471a1 · inbound

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning cites this paper.

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:42:49.502198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T23:38:38.127345Z digest=sha256:8d381db8b3d393b418a85d6411867faec10e1400631ec2cc636bc9caa9ba5482

Observation bc748230-e09d-4318-9bc3-3bbde11e7e04 · inbound

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation cites this paper.

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:37:37.317342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T13:45:13.320426Z digest=sha256:2d4fae752ac1006059db4408ebc499f20bbde9faa9d8dbd905ed83e004c984e3

Observation c7849f53-caad-47c5-a841-a73812005f57 · inbound

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation cites this paper.

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:24:38.477330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T11:15:21.946819Z digest=sha256:87c80254271917e3b3f9a6d61391d797949da4972a73d246bebe569c930e78cd

Observation 4a1fe570-b980-444e-aeb7-1dcea8de1d82 · inbound

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal cites this paper.

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T09:07:47.865564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T10:32:57.295159Z digest=sha256:a876c218c089a90dbc9fd6a521bef529d09e6d34589dad481e8f4345c5c6fa8d

Observation e72b5fd4-22cc-4ebe-ac0b-331a1ec72b46 · inbound

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs cites this paper.

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:37:56.589387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:57:14.328157Z digest=sha256:52439ffe1b26c2189a7f8053362b7e6a7260b7ac74e18834712e305d8f98b8ad

Observation 63f1c808-701f-4288-8d42-b8a70539c942 · inbound

Probe-and-Refine Tuning of Repository Guidance for Coding Agents cites this paper.

Probe-and-Refine Tuning of Repository Guidance for Coding Agents Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:29:35.754233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T15:58:40.881243Z digest=sha256:1dfd02aae1e966ecf2350568d7914474c6b822e5d4c231bdeb55e7d93b00d9a8

Observation 88be5b9d-b074-475b-99f5-f940d587f5a2 · inbound

Channel Location Constrains the Auditability of Subliminal Learning cites this paper.

Channel Location Constrains the Auditability of Subliminal Learning Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:19:44.436777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T11:52:03.948568Z digest=sha256:cc99704acf183fdd2a4db9ef884f46676d625455bbce81906ddca7625af587c3

Observation eb938279-88f7-4753-a87e-8ff6ecc6d4d6 · inbound

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology cites this paper.

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:07:08.140193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-07-02T15:57:48.589980Z digest=sha256:e3a5c8a4923c51a930b1079d49afce0ef1600f8f9fc73b936538bfda4125cda4

Observation 6c35d487-4c19-40ce-b04f-fc67b8a7cce2 · inbound

Unsupervised Features Mining via Activation Geometry cites this paper.

Unsupervised Features Mining via Activation Geometry Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T20:51:42.052454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T20:51:42.052454Z digest=sha256:6c14d09cb0c9f8d2d1b2334136e55f0df2bfb2c4a7b6b2eda8f76fa59ec25bd2

Observation 7fa3c2e8-c11c-4269-8911-963ae569502c · inbound

Emergent Misalignment Recruits a Pre-existing Persona Subspace cites this paper.

Emergent Misalignment Recruits a Pre-existing Persona Subspace Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-01T07:46:09.323606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:46:09.323606Z digest=sha256:0052a7f092860cf8aa11dcf87ef772d1754b47ccea45599e3fde231358aa5ee4

Observation b5dec608-6606-45fc-abbd-a41cec86ab9a · inbound

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models cites this paper.

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T00:38:41.279652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:38:41.279652Z digest=sha256:e0b88f9f22d97ac5c437a32063435af38bf6d04886d223eeed7f9ef312311c73

Observation 2c681241-f1ae-4436-add0-4fa973d9c1e6 · inbound

Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems cites this paper.

Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:13:34.537894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:13:34.537894Z digest=sha256:040205e20a58ee18ac3ec7550b6bcbcd9618bfe979138fc12688f77df7e758e1