Pith. sign in

Paper Citation Record · LEDGER

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models

As of 21 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2506.12217.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12217 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T01:03:43.999176Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:32.686738Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1629fedd-c51a-443c-8c86-a9e44716013d · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.561365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.561365Z digest=sha256:58df3d4a0624bf790050f1512285cb423efd18b59cf48079590172e7426da8a8

Observation c3f47b45-1cf7-4040-b36c-88f44fb55353 · outbound

This paper cites Self-reasoning language models: Unfold hidden reasoning chains with few reasoning catalyst,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-reasoning language models: Unfold hidden reasoning chains with few reasoning catalyst,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:50.091457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:38.644078Z digest=sha256:1730834dcd2bd21ad3a6ec366735fd0ac2a1cc4c6b92c2c1948ba3552544d89c

Observation d39c0232-4edf-4f1f-93f0-77c3ec0e014d · outbound

This paper cites Reinforcement learning with verifiable rewards: Grpo’s effective loss, dynamics, and success amplification,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement learning with verifiable rewards: Grpo’s effective loss, dynamics, and success amplification,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.746258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.746258Z digest=sha256:b8ade82cfd542f5683e08aae96956851799ac08597225c03d1f2bc2330dff26d

Observation 72ccaf44-fc65-488a-a1cd-10b9b7ca1794 · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.824120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.824120Z digest=sha256:f187424b38239bb4332b3b57c2d430497582dfd5152d2bc5f3f6da6ce7d133cd

Observation 95f02528-e180-46f1-b36e-a332307da8d4 · outbound

This paper cites Reasoning beyond limits: Advances and open problems for llms,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reasoning beyond limits: Advances and open problems for llms,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.946446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.946446Z digest=sha256:5d6c4544bf821fb42995361fe01d7b1a99ef574aefb7af40d5b0aefa4dc0122a

Observation 45d137f3-ca87-4e2f-b01a-e084297c6e34 · outbound

This paper cites Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.063087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.063087Z digest=sha256:6ea4968997072588a1c10e9cc01581fa7f9426fe5c52593bd1b9c424b8277e53

Observation 6ce3f4ec-e201-4c21-8cfa-de0702166cd9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.150691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.150691Z digest=sha256:bcdd150ec3e4523c88b323dc6de31e054714be3193ab245e57123b297fb288ed

Observation 00266b23-2c47-4af4-a0f7-ac7c0164fc40 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Understanding R1-Zero-Like Training: A Critical Perspective

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.248657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.248657Z digest=sha256:3be12a623a74d8699d828c6f9064bc81c73e9aa481b07e95999b1c149e01f175

Observation fec69ab1-664a-4a23-b22f-5e9923cbb2d1 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.348782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.348782Z digest=sha256:1047174f65ea3c35b296c2c71b9dab754e89450fe1ccc20b37220db2b4c46601

Observation 39da37bb-30e3-44ac-a74b-d2899675fa2e · outbound

This paper cites TTRL: Test-Time Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models TTRL: Test-Time Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.439450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.439450Z digest=sha256:8503be51cca1a4d7797bc5daea684f9f408c50de0fe3a72af8e330c2466439ba

Observation 6e48c6ec-e596-4e7a-a061-ee7584b866ec · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.555965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.555965Z digest=sha256:8944b535f7ae10e3d153dbea8877d02023b85dc043730db947cc81ed0a593e17

Observation 5012415e-37d7-479a-b840-2af01d93001e · outbound

This paper cites Self-Reflection Makes Large Language Models Safer, Less Biased, and Ideologically Neutral.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-Reflection Makes Large Language Models Safer, Less Biased, and Ideologically Neutral

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.646395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.646395Z digest=sha256:dd2575c3f56884c8d6220346161ec099d2fd4a82d8b7ce03eee1d2d85147a211

Observation 3a583017-4b0b-4ce9-a51f-739836db59d9 · outbound

This paper cites Dynamic early exit in reasoning models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Dynamic early exit in reasoning models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.763786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.763786Z digest=sha256:581c9e28df8fa3b5c3ec8bd51e4de2f71792a02bedf47e4ab199e027d822ae14

Observation 0ff7c1eb-0746-4af3-a970-991237b9ad2b · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.875381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.875381Z digest=sha256:036f26de8110ca69d9e9181e047285b340869e47aad903935bf1441f498bf98d

Observation e75301d6-8fc0-4c2b-8fd8-258fd41755a5 · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.949009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.949009Z digest=sha256:bc0a6b462b8ae42857bbe3e8ae67a5301f49c30dd02e3d63d3820b56dfc1f2c6

Observation 1c3c4d59-7537-49e7-87a4-48675c718d1f · outbound

This paper cites Steering llama 2 via contrastive activation addition,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Steering llama 2 via contrastive activation addition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.915796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:40.075252Z digest=sha256:51c796540ccf6bed40625f49ba110efb19ed1e21b062963c3ebf17ff1115d71f

Observation dfcd3ad3-4410-464d-bd53-7cafe43242bd · outbound

This paper cites Generating Wikipedia by Summarizing Long Sequences.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Generating Wikipedia by Summarizing Long Sequences

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.150110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.150110Z digest=sha256:4abcc197cd6bccc8935b4acb98ccef7664ccf5533e4555365bc629d4162c17a1

Observation d7c5a580-fc9f-4a8e-afd2-9006c9ee36f6 · outbound

This paper cites Beyond accuracy: Evaluating the reasoning behavior of large language models - a survey,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Beyond accuracy: Evaluating the reasoning behavior of large language models - a survey,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.713626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:40.258094Z digest=sha256:3312303fc1e6bdbe8ac1be09ee91c45ffe6c7d78eeb9398ed5ac7d5cd7fdb08d

Observation 6edd9d22-89a5-4dc0-9c6f-af10c5a203a7 · outbound

This paper cites Oat: A research-friendly framework for llm online alignment,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Oat: A research-friendly framework for llm online alignment,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.490397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:40.356585Z digest=sha256:f85a00f23377f532a3f1038467612a63bdab094bed7414f35f9b42efe714c383

Observation aa276d97-fc20-4b0b-9055-911fb8bdbae9 · outbound

This paper cites Self- consistency improves chain of thought reasoning in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self- consistency improves chain of thought reasoning in language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.324127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:40.427956Z digest=sha256:5cba91ffa4e3c3f9746f567e4d7863f7e38e348f143f727bbe946b8bf655c983

Observation f3ca3234-ffce-4943-a683-0ebd222aaa98 · outbound

This paper cites X-reasoner: Towards generalizable reasoning across modalities and domains,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models X-reasoner: Towards generalizable reasoning across modalities and domains,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.145880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:40.536878Z digest=sha256:d405d61a614606f0fdfda4e37971802e54b41a55281404c28becd531de5904af

Observation 271c3ef1-143a-4feb-a781-27cf896f307c · outbound

This paper cites Reinforcement Learning Enhanced LLMs: A Survey.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement Learning Enhanced LLMs: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.631750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.631750Z digest=sha256:78682b943b2e65292517e3c5a9573193bdcb05c7b15f5a9f80c0ebdf09f7a9fb

Observation a66ec951-743b-4545-99c4-8e4e8a841cbf · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.693792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.693792Z digest=sha256:41fbb4375e20a06a49058b645a4aabe81c00f0983287331767c2d58b91afd570

Observation f9f122bc-638e-4812-94d2-a741f75f391c · outbound

This paper cites When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.835799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.835799Z digest=sha256:c3f3588a14a85135506535c6c5b1710d934d0e0ffc83793ec82d09217698f54f

Observation d7ed41c2-594d-4bc7-8560-4063ddfa933d · outbound

This paper cites Demystifyinglongchain-of-thoughtreasoninginLLMs,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Demystifyinglongchain-of-thoughtreasoninginLLMs,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.981885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:40.918114Z digest=sha256:b35e0828f986a75d0b4763ab7eaf0d82640d1f66956c72dccefe2551d63e56d4

Observation 9646293b-ee21-4f86-9fc0-d89a56d2de0f · outbound

This paper cites OpenAI o1 System Card.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models OpenAI o1 System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.000635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.000635Z digest=sha256:9dc530359c5423e1ed3c846b943fce635162e7b0f87b90d38429d9034cdfe7a6

Observation 469303ad-d1e6-4663-a969-09b9a3aad043 · outbound

This paper cites 2 OLMo 2 Furious.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models 2 OLMo 2 Furious

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.101622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.101622Z digest=sha256:f383d5cf2d03e3084a82bdc24cf4e0b39160c1cbc653ed50bf45182588d34807

Observation a6bebd1d-2878-441b-b47a-317a95e26e40 · outbound

This paper cites Qwen3technical report,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Qwen3technical report,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.803809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:41.187541Z digest=sha256:eb7c52263aec65922d472e8df7e0e7eadb3b947fe478cf1541c60613e18772ba

Observation 7c38a833-c42f-4bdf-b9e3-6d709d13d656 · outbound

This paper cites Measuring mathematical problem solving with the math dataset,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Measuring mathematical problem solving with the math dataset,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.604954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:41.281529Z digest=sha256:45b3225643d71847d0b040277e662e45e0a7b28cef79e40d6b77bcf0f393c9d8

Observation c3b0ce6b-1dcd-4094-a847-9ac37c243daf · outbound

This paper cites Qwen2.5 Technical Report.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Qwen2.5 Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.367515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.367515Z digest=sha256:210428ad4cf489f7075a18ed701feb6834625a9fa3f454b8ce854726f5879746

Observation 5f2a934e-8188-4b61-b082-cfc0a740d4f2 · outbound

This paper cites Umap: Uniform manifold approximation and projection,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Umap: Uniform manifold approximation and projection,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.402203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:41.455519Z digest=sha256:d67e82db2ec6b03326a4ce790cfbdf04a8644e8348e6ebb9a42666bc916e2c9b

Observation c48b7e33-6880-4a41-9bcc-4e2415ae0261 · outbound

This paper cites Refusal in language models is mediated by a single direction,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Refusal in language models is mediated by a single direction,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.223985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:41.581041Z digest=sha256:ff7be5e44c248a4704668b8d70ecfb4201ef6afce18232f015679a9f57954909

Observation aa3cfd68-4461-42d7-9c7b-067819b1df5e · outbound

This paper cites AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.633267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.633267Z digest=sha256:644d232736cc7ca78ead965474edc395f1c36e5312f77168f6b9c22b2e3dec1f

Observation 574728a9-8514-40e5-808d-9ebd8bffb441 · outbound

This paper cites GPQA: A graduate-level google-proof q&a benchmark,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models GPQA: A graduate-level google-proof q&a benchmark,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.007208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:41.688003Z digest=sha256:f25dd956862f3a4277e5b969aee41d6fb8997b8cf8ee14cb448adc25ac61489d

Observation 663e3187-2685-436c-8ec6-a625da5ab9ca · outbound

This paper cites The Llama 3 Herd of Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models The Llama 3 Herd of Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.772289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.772289Z digest=sha256:4b3429ff63beaf633a14c32efa03643525d14a06ccee43d22071ec8e46db0594

Observation 176a923f-3bc6-4556-bb92-d9d3e3b7a73e · outbound

This paper cites s1: Simple test-time scaling,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models s1: Simple test-time scaling,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.843540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:41.883084Z digest=sha256:53173514131d04da0dfbe05ef6ba8231c84c07145e68e880cf50ba254e6075fa

Observation 455bfbd3-f1f3-4569-9fb0-13256cdc8c65 · outbound

This paper cites Discovering latent knowledge in language models without supervision,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Discovering latent knowledge in language models without supervision,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.687530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:41.999426Z digest=sha256:2045a2e81d62e4ac71b00fba2bf3b22ea0f10881d7181007c8ffb87c2140e887

Observation e080e39d-6fdc-46bd-b468-be482af58d7d · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.097166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.097166Z digest=sha256:bcc52f706c2585d959ba2f02caea5e384fe938957de24567fbba165b6d1c474e

Observation 066fae06-1e46-4c3e-943f-2ef5c7231f6c · outbound

This paper cites A Language Model's Guide Through Latent Space.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models A Language Model's Guide Through Latent Space

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.165517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.165517Z digest=sha256:c333c41af226c199b5616b50a4538ea30d03a49b06774cb1381fccd1f7da1c2f

Observation ccda8ba5-f1c8-4351-b338-06c32a8b5839 · outbound

This paper cites Improving Activation Steering in Language Models with Mean-Centring.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.231617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.231617Z digest=sha256:9be2a3b994a2b714ad9d6d711665bbd6ea3b915855ee9f6a6d7f8053987cd1bc

Observation 4b8fa7cc-b175-46cf-a4bc-2e07327b259b · outbound

This paper cites Finding alignments between interpretable causal variables and distributed neural representations,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Finding alignments between interpretable causal variables and distributed neural representations,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.540585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.293333Z digest=sha256:51ac2d42d1a1265b9097ce35a3722af71149d2d763f3da3ea4b8d03a72a1bf7c

Observation 75b62a26-f9ce-45a9-bf6d-25a79750aa6c · outbound

This paper cites Generative agents: Interactivesimulacraofhumanbehavior,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Generative agents: Interactivesimulacraofhumanbehavior,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.341340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.349684Z digest=sha256:990f78817c90a6070a8c435dd5921c968dd64aff62dfa562080d62a993cfda95

Observation 2e2ee699-326c-43f4-9ce6-e929d40d1b2a · outbound

This paper cites Man is to computer programmer as woman is to homemaker? debiasing word embeddings,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Man is to computer programmer as woman is to homemaker? debiasing word embeddings,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.140600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.442869Z digest=sha256:7bf1f6274eb5bf0c8d2183d028e18c0fd7dec2227d32501e717261eff8c86ab8

Observation 5646bbed-0a47-48ee-92c0-a8de0ca8edb5 · outbound

This paper cites Sparseautoencodersfindhighlyinter- pretable features in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Sparseautoencodersfindhighlyinter- pretable features in language models,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.980720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.512310Z digest=sha256:d892b9fd5353ac8975a894641412ee12d605b93c6bcb00ec5b85ebfab7fb23e6

Observation f7461889-ce63-41fe-89a9-fbd38f2db7d1 · outbound

This paper cites Gold doesn‘t always glitter: Spectral removal of linear and nonlinear guardedattributeinformation,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Gold doesn‘t always glitter: Spectral removal of linear and nonlinear guardedattributeinformation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.832967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.632456Z digest=sha256:c501d5a1d8707f51d52b512671bce619841e07d209ae70a906554911d385eba0

Observation ee86d914-0601-459b-b18b-37f5071d3aff · outbound

This paper cites LEACE: Perfect linear concept erasure in closed form,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models LEACE: Perfect linear concept erasure in closed form,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.659998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.748817Z digest=sha256:0ea770c602b9db04edf998bf1c3f8d36de969b4b98434a9b807658c97316303f

Observation 4876e0ea-6e93-4259-9ab9-3a14a1faeb66 · outbound

This paper cites Monitoring latent world states in language models with proposi- tional probes,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Monitoring latent world states in language models with proposi- tional probes,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.506000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.820828Z digest=sha256:e7c917c5ee1dc2452c3d60be4d1b1408470d873681eef0b1099e024565b58482

Observation 14c8a776-3a4d-498f-8938-059650e3aa86 · outbound

This paper cites Let’s verify step by step,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Let’s verify step by step,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.336643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.869100Z digest=sha256:0241e27f6dc6a1423c3398d5e80ea2e892b9f51525b7ed158d45a5a3d4a9f13b

Observation 45964bdf-c1b5-4127-be86-22247b70986f · outbound

This paper cites Self-refine: Iterative refinement with self-feedback,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-refine: Iterative refinement with self-feedback,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.194754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:42.995770Z digest=sha256:00e7bf4bfc2b042bd365976642bc58ca6ee49899ddffc8b0bfd6efb9975a8d07

Observation 4f77234f-d40c-4b67-ac9a-698fee7368e6 · outbound

This paper cites Fine-tuning with divergent chains of thought boosts reasoning through self-correction in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Fine-tuning with divergent chains of thought boosts reasoning through self-correction in language models,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.041393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:43.095901Z digest=sha256:ca4a9665f7599910910d8a9bfb1f2ed6140511c965431347d05587a1463df977

Observation b0fac060-c3ff-45bf-be9d-73d9cfe73d47 · outbound

This paper cites STar: Bootstrapping reasoning with reasoning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models STar: Bootstrapping reasoning with reasoning,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.858720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:43.195392Z digest=sha256:3cacb23b0e7fae90e1edccd0f67fe81e7fbc176d6526858c364bf69fc4b8001c

Observation 2912de11-1a0d-4b74-84d1-31f3dffcfaae · outbound

This paper cites Let’s verify step by step,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Let’s verify step by step,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.699379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:43.254472Z digest=sha256:19ef78582807afb4c1f928cec7e60aaf20bef9c500978bcb2050d9fc601818e7

Observation b5856e66-e3bf-4112-b2c0-a838138b03f0 · outbound

This paper cites LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.321333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.321333Z digest=sha256:564d5aed33a749ee151c02b1814cee879f79e22846ad69a9fc5592ea5dbefb12

Observation 1af91da7-bec5-4818-8733-323b5f4b5934 · outbound

This paper cites Shorterbetter: Guidingreasoningmodelstofindoptimalinferencelengthforefficient reasoning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Shorterbetter: Guidingreasoningmodelstofindoptimalinferencelengthforefficient reasoning,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.439679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.439679Z digest=sha256:3850cc6b0b7d407ddb1cacb605736998c03f721e92024d53213c9b9b09b10de1

Observation ee2f8254-4724-4d15-b0f4-e5a60cf4666e · outbound

This paper cites Unlocking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Unlocking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.472186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:43.510697Z digest=sha256:10b43ddcff527857beac39cf3226bcc9d2af31b30eeb9ca32c6388bd5e922f22

Observation 1e93143a-2574-4054-be9a-79baf7e88ac1 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.598058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.598058Z digest=sha256:0083d891a459043a02731c7291c2d34404d179dbeb1e54a70c4596b7486b3b98

Observation c86d77b5-78b6-4952-baaf-a77e3547715d · outbound

This paper cites Rethinking Reflection in Pre-Training.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Rethinking Reflection in Pre-Training

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.678277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.678277Z digest=sha256:5d0c4823823ffbf3a7190c4c393675f7b515da0633ea36d4de93627f8fc3dcc3

Observation ab5573a1-46ee-49f2-b945-5d27a0a59f08 · outbound

This paper cites Reflexion: Languageagentswithverbal reinforcement learning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reflexion: Languageagentswithverbal reinforcement learning,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.213263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:43.774554Z digest=sha256:26000fb81e2997e7bd1a2737f34314f8a70ff40b3b9493dd3a3363cc0ecf8d26

Observation 3fa9b162-1ea8-45d9-b1a8-faa954e327d7 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.838361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.838361Z digest=sha256:3fb43bc1ce053d15d2f57eb0c98b90788a4360876645f51e68ad8c773eb2894b

Observation 33193c25-3851-45af-84ae-f7a83be9ac32 · outbound

This paper cites There may not be aha moment in r1-zero-like training — a pilot study.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models There may not be aha moment in r1-zero-like training — a pilot study

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.036245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T01:03:43.999176Z digest=sha256:b4c5b238860744228fd3bdad4fa01b1a60eebbeec80aa86b381994c63ad3592f

Pith citing papers

Observation 09e6d2ee-3503-4e81-8bf6-3281c83c64a6 · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges From Emergence to Control: Probing and Modulating Self-Reflection in Language Models

Reference 201

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:32.686738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:32.686738Z digest=sha256:c5e389534154e4a3b9565a1bdf8604affeae8652736f9f1a8c703a8802652156