Pith. sign in

Paper Citation Record · LEDGER

Learning by Distilling Context

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 43 inbound Pith citation observations for arXiv:2209.15189.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2209.15189 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:13:38.311449Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

7
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6315a156-3a7e-4e8f-8087-f5ed01a82869 · inbound

Large Language Models Can Self-Improve cites this paper.

Large Language Models Can Self-Improve Learning by Distilling Context

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:00:48.263694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T17:00:48.167441Z digest=sha256:e1b6e54df3924958ee72c02b551eace8c9231f25238e1d077f38de5a005009fd

Observation 8a76589d-0a1a-49f6-b3ab-3bce8130e323 · inbound

The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions cites this paper.

The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions Learning by Distilling Context

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:59:30.789754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T10:59:30.728091Z digest=sha256:54c1c0dcdad182f60c53941f11c9d189469198f18e635be964457b845ee7397c

Observation 7c63712a-c751-46c4-8876-d99ce7a1cfde · inbound

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models cites this paper.

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models Learning by Distilling Context

Reference 297

Resolution
verified exact
arxiv_id, observed 2026-05-17T14:43:29.847490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-17T14:43:29.496457Z digest=sha256:ab70b3750f1697cffed237147877818e03ba348a5100ac5a5830037a0bf0ae17

Observation d0e9db7c-3374-4f60-b279-0a9bc212c816 · inbound

Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents cites this paper.

Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents Learning by Distilling Context

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-08T14:13:38.311449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:13:38.311449Z digest=sha256:6bb5d3a72d20d45b0a17e90db5f6e9f6b43f7a7267db43ddf3fc67c264a14ce7

Observation 9b8246c5-08e8-412e-bdcf-56627b82295c · inbound

Cartridges: Lightweight and general-purpose long context representations via self-study cites this paper.

Cartridges: Lightweight and general-purpose long context representations via self-study Learning by Distilling Context

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T06:04:34.304670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:04:34.304670Z digest=sha256:bea1d25ad3624c09968101d5c78aae6c92e033048f560ebee6c54d9ef7b4c70e

Observation c544393c-4aa6-4ab5-8fa6-30f90d333d8f · inbound

Essential-Web v1.0: 24T tokens of organized web data cites this paper.

Essential-Web v1.0: 24T tokens of organized web data Learning by Distilling Context

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:46.861099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:24:46.861099Z digest=sha256:df279fc2d1bf2e172c46f109a76c76c64755f4f1ff7d673b455e372c4d38c1b3

Observation fda6bd76-14b3-4ca1-ba50-712198c84b62 · inbound

AgentDistill: Training-Free Agent Distillation with Generalizable MCP Boxes cites this paper.

AgentDistill: Training-Free Agent Distillation with Generalizable MCP Boxes Learning by Distilling Context

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:38.518193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:38.518193Z digest=sha256:87098ad16225bbb5baf28dd092af8090eac8f4d1191216d992c1c9f4fb9fc5fc

Observation 50903d1d-fcc2-4a71-87ca-4b368a71f632 · inbound

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation cites this paper.

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation Learning by Distilling Context

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:06:01.325654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T16:48:53.092395Z digest=sha256:7b9eb5f01ef300146201111259c0b90abe75c1e18fb1e389e9f74400849c99be

Observation 6b05a7e0-2cd5-41ea-9ae2-83ab31691c93 · inbound

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning cites this paper.

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning Learning by Distilling Context

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:30:53.811574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:28:58.515666Z digest=sha256:fd44d91bd23177b9ce04fb0d412a7bb094c0cc963782fba2ed3b9bae381bdad5

Observation 390f6538-ff00-4411-b80c-8a8ba392cbc8 · inbound

Tuning Qwen2.5-VL to Improve Its Web Interaction Skills cites this paper.

Tuning Qwen2.5-VL to Improve Its Web Interaction Skills Learning by Distilling Context

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:51:36.168174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:51:00.660022Z digest=sha256:42ea9cd8827fcea1df12dd080d95d7bfed885890910ea738f51ada12c9049fbb

Observation bc112bd6-1564-4118-b3f7-1c480e323da7 · inbound

Near-Future Policy Optimization cites this paper.

Near-Future Policy Optimization Learning by Distilling Context

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:54:48.621807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T00:51:36.580600Z digest=sha256:181331c37251e324fc74bbe05275e89a9afd51e0fa12971f66bcbc6ed8ad432f

Observation 23e15da8-ae9f-4e89-a077-959b1c3a9c0f · inbound

Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts cites this paper.

Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts Learning by Distilling Context

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:50:57.396716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T02:11:19.295354Z digest=sha256:4a2f1cd059cf303e079d23d0497b2c07b97a5be9f9a255463985b312da59baff

Observation 0d3e8055-9993-4172-b246-d8d7ee956d9c · inbound

CoDistill-GRPO: A Co-Distillation Recipe for Efficient Group Relative Policy Optimization cites this paper.

CoDistill-GRPO: A Co-Distillation Recipe for Efficient Group Relative Policy Optimization Learning by Distilling Context

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:31:26.461256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T00:59:44.364491Z digest=sha256:b538ad6a10cd128bb86566329d74f9dbb34729db1b7ced04427d9839b4aeb0e4

Observation 2f30f395-0463-46f1-8409-6420710e6aed · inbound

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why cites this paper.

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why Learning by Distilling Context

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:11:27.211928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T04:28:58.679078Z digest=sha256:5903c950a7b37d6eba6eace48ffb3c33c29e897ab5ece732b67d3d498c75009b

Observation fe2f5de1-d1ce-4a95-88cd-496aadba5cd3 · inbound

From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation cites this paper.

From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation Learning by Distilling Context

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:27:02.004969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:24:34.093541Z digest=sha256:88803dc155c86356a14b90d9dc121665c080bf476abd60e5effaa28410be9810

Observation 8d283293-324b-431b-a553-37de5b1c0d88 · inbound

VSPO: Vector-Steered Policy Optimization for Behavioral Control cites this paper.

VSPO: Vector-Steered Policy Optimization for Behavioral Control Learning by Distilling Context

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:43:44.111000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T19:39:55.294398Z digest=sha256:c69e13e5c575bb46b1e59a470cd4ed64462fca56388666db3ad5ab75147fa4df

Observation f2e52020-fd83-41d3-9150-d734f9a25543 · inbound

Self-Supervised On-Policy Distillation for Reasoning Language Models cites this paper.

Self-Supervised On-Policy Distillation for Reasoning Language Models Learning by Distilling Context

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:43:21.837694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T14:42:55.368104Z digest=sha256:dfc5230ebc7c1a9d065aa9799e7a027db3cf312d01f18ada85e74fe1334ce5fa

Observation c684b1c3-bc07-4859-bff8-e44a0c0ec5e9 · inbound

Context Memorization for Efficient Long Context Generation cites this paper.

Context Memorization for Efficient Long Context Generation Learning by Distilling Context

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:43:12.621794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:39:09.720412Z digest=sha256:50a8c2c5a66e4de568ccb6c88a203fb7c14c5e5f7e7b094bc3b30d43c1533b95

Observation aead6aed-ce7e-46bc-9f79-d599385a5c80 · inbound

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs cites this paper.

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs Learning by Distilling Context

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:09:51.844225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T08:05:44.358256Z digest=sha256:b07131dfda0f940bb18e81ac7f59fccdc815f0b9b2103ff745e604de5ddb7d11

Observation c02f2b1f-98f7-48ae-a4be-57807067c3f1 · inbound

Tailoring Teaching to Aptitude: Direction-Adaptive Self-Distillation for LLM Reasoning cites this paper.

Tailoring Teaching to Aptitude: Direction-Adaptive Self-Distillation for LLM Reasoning Learning by Distilling Context

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:11:17.794411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T08:06:30.862911Z digest=sha256:449368db1e7830bd0b744a66b62554fd037c139a807989e7f2b307f665fb54d0

Observation e0a87535-1268-4e94-a7ce-4f26e4217fc5 · inbound

Do Language Models Need Sleep? Offline Recurrence for Improved Online Inference cites this paper.

Do Language Models Need Sleep? Offline Recurrence for Improved Online Inference Learning by Distilling Context

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:43:59.454603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T21:37:51.638904Z digest=sha256:0af13b385747c450382e93de2fc50f5307714ab3a67bb1cf0e8bf103cd52e43c

Observation 9f489a76-dc34-4a8f-8804-8a86246c3e36 · inbound

ThinkSwitch: Context Distillation with LoRA and Weight Interpolation for Specific-Purpose Reasoning Tasks cites this paper.

ThinkSwitch: Context Distillation with LoRA and Weight Interpolation for Specific-Purpose Reasoning Tasks Learning by Distilling Context

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T18:02:27.323880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T17:52:59.979543Z digest=sha256:1f758de2702a2486fdd6b281356516af4df1a29441a22974e4282213bff67d94

Observation a27e4b2b-54bf-489c-97b7-9ffcc22ee17e · inbound

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents cites this paper.

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents Learning by Distilling Context

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:56:47.807293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T06:29:11.398007Z digest=sha256:754cd44fb61ac8763f1fcb653d75ffc706da93f3b9e9527afe740b0b7d186484

Observation df9f27f8-938f-4796-91dc-a98b8b99672d · inbound

Amortizing Federated Adaptation: Hypernetwork Driven LoRA for Personalized Foundation Models cites this paper.

Amortizing Federated Adaptation: Hypernetwork Driven LoRA for Personalized Foundation Models Learning by Distilling Context

Reference 76

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:16:58.682443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T01:27:04.241484Z digest=sha256:97f0a83f413a471f4fba448df8ee1ef6951e6e86ed76d0c7d42cd83389888e12

Observation 2a5df5e2-4795-432a-a14d-fe02e61050d1 · inbound

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning cites this paper.

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning Learning by Distilling Context

Reference 205

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:17:48.415598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T10:28:11.440915Z digest=sha256:4d8e33bd2de79e167c4a84559e40e246cf0c33268c20b168507d0c5426dda73e

Observation e8b6433a-3038-4a7e-a9f2-ffffa1b77373 · inbound

Doc-to-Atom: Learning to Compile and Compose Memory Atoms cites this paper.

Doc-to-Atom: Learning to Compile and Compose Memory Atoms Learning by Distilling Context

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:37:57.065542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T09:53:35.191873Z digest=sha256:5fd77c0d390f8f58e92689c1140847291c801aaad82ac4ae05998d985151b73c

Observation 59ec0084-5f5e-4a98-b02b-9789f1e586ad · inbound

PRISMR: Overcoming Parse Collapse in Multimodal Listwise Ranking via Parameterized Representation Internalization cites this paper.

PRISMR: Overcoming Parse Collapse in Multimodal Listwise Ranking via Parameterized Representation Internalization Learning by Distilling Context

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:38:29.012031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T07:00:17.978103Z digest=sha256:07de89d659412a22fbfe4a07fde34b3407ce7e0ee5ee859dc059dbc00d0f2395

Observation 1e7fea36-569e-4851-a290-89d780eaa74c · inbound

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis cites this paper.

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis Learning by Distilling Context

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T13:53:50.979339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:53:50.979339Z digest=sha256:8457ecc147921983c6b71db9b473047ce1597b83b5dd9ec9796127e2ba2e18c1

Observation 5dd966a1-a03d-4b16-8ed1-aaccc18458a8 · inbound

HMARS: A Hierarchical Multi-Agent Memory System for Long-Context Reasoning cites this paper.

HMARS: A Hierarchical Multi-Agent Memory System for Long-Context Reasoning Learning by Distilling Context

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:34:37.890032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-30T11:30:52.871762Z digest=sha256:784cd651538305b234c317dba14ec57cbd21d2706d06090ee3ab1adf9f0dc890

Observation 6ad6097b-b6b7-458b-870b-41f3ce6d15c3 · inbound

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization cites this paper.

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization Learning by Distilling Context

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T12:05:43.607060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-01T02:32:19.425550Z digest=sha256:845f4f2c1bcaa2602d462c889dc48b797ec5e3c2842a878edc5493b22c5cf9f0

Observation c08528f8-3569-4cad-b504-2a86a7fb4596 · inbound

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation cites this paper.

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation Learning by Distilling Context

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:36:56.170129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-02T12:27:59.332432Z digest=sha256:98c97f6a5f148846e8751bfa552582b15f35fa5e74a5ada61a0e4a2c4d93f3f7

Observation ccca42a2-76c5-4d5d-b4ce-89280ded27cd · inbound

Epistemic Goggles: A Pretrained Module that Induces an Epistemic Frame via Gradient Editing cites this paper.

Epistemic Goggles: A Pretrained Module that Induces an Epistemic Frame via Gradient Editing Learning by Distilling Context

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:28:31.104704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-03T14:25:39.401131Z digest=sha256:565104e572786098fb659f60096a953545316567bca4bfa3e319502dc9007e97

Observation 705eea4c-835b-4c44-8dac-9d3fa8fca4d3 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents Learning by Distilling Context

Reference 107

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.163623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:d9ba50a97c3f295537bce22e9baab28e49aa805c4f22d3e268f86aecf4d9841e

Observation 8f316d1a-e5dd-451f-915c-2364c8772c32 · inbound

Can a Language Model Learn Facts Continually in Its Weights? cites this paper.

Can a Language Model Learn Facts Continually in Its Weights? Learning by Distilling Context

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T07:36:43.499258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:36:43.499258Z digest=sha256:a646bf397ea0a6b6308278f2d365d57070897b124f87cd55af29effc23b3e9b5

Observation 7f4310bc-91e5-42bf-8c39-a8b170fc4602 · inbound

Sample-Efficient Learning from Agent Experience cites this paper.

Sample-Efficient Learning from Agent Experience Learning by Distilling Context

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T08:43:21.932131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:43:21.932131Z digest=sha256:8c2c47b7b34d1c2a1837d569c07cf82456b336dd853675aa103c1a1311d36888

Observation b9d82d3a-535f-425f-8fd3-486c7f062d7e · inbound

Do Modules Stay in Their Lane? Role Drift in Compound LLM Systems cites this paper.

Do Modules Stay in Their Lane? Role Drift in Compound LLM Systems Learning by Distilling Context

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T08:17:26.367813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:17:26.367813Z digest=sha256:29ba52704c82e6a115517ed819cca57262eca239661e5b266bbb94d646d3cf3b

Observation e6c8895e-7140-4b0d-9c91-a653f644cb73 · inbound

Listen, Do Not Copy: Internalizing Audio-Grounded Scaffold Context for Robust Omni-Model Speech Understanding cites this paper.

Listen, Do Not Copy: Internalizing Audio-Grounded Scaffold Context for Robust Omni-Model Speech Understanding Learning by Distilling Context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T06:19:33.044773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:19:33.044773Z digest=sha256:42e827594dd969b6e426f498344604fba84a1433938a3e1a2303ee5ba6f32ef5

Observation 7bad19ac-b647-4b4c-ab95-a94ab1ce032d · inbound

Masked Distillation: Internalizing the Chain-of-Thought in Language Models cites this paper.

Masked Distillation: Internalizing the Chain-of-Thought in Language Models Learning by Distilling Context

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T10:49:24.977221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:49:24.977221Z digest=sha256:de04519c3b7bf1c6f6e501b3c2077f5317d62a59623a535b5961e94d27ce88ce

Observation f34e8e4d-f88c-43da-a22c-1dd43c616bee · inbound

Flux-OPD: On-Policy Distillation with Evolving Contexts cites this paper.

Flux-OPD: On-Policy Distillation with Evolving Contexts Learning by Distilling Context

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T20:13:02.033825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T20:13:02.033825Z digest=sha256:a2fd4041a2eeb12c0a0c79f34f91ea710c1fb4ca5488db855e338ad196f94a90

Observation d702b95c-08f7-4a5f-ae0c-2adbb5ce5d7b · inbound

Learning What to Remember: Test-Time Training via Context Distillation cites this paper.

Learning What to Remember: Test-Time Training via Context Distillation Learning by Distilling Context

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T23:14:10.159095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:14:10.159095Z digest=sha256:54abe25f5fbcd0d2d17a3dc21701640160ee241f422db86fb36de3a865d2e9b6

Observation d2d95894-4d46-4742-b75e-b27fa2000d45 · inbound

Instruction-Conditioned Exploration for Reinforcement Learning with Self-Distillation to an Unconditioned Policy cites this paper.

Instruction-Conditioned Exploration for Reinforcement Learning with Self-Distillation to an Unconditioned Policy Learning by Distilling Context

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:16.462866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:15:16.462866Z digest=sha256:0d7631c8f574d4c1bd40fb3a8424449dec1630e019a09f7b6a12c68ee90a4f55

Observation ff8d01c5-2bd5-42d1-9e20-50c0d9be4d9a · inbound

Agentic Reinforcement Learning with Self-Distilled Reward Shaping cites this paper.

Agentic Reinforcement Learning with Self-Distilled Reward Shaping Learning by Distilling Context

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T23:15:50.848819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:15:50.848819Z digest=sha256:e7a0f77eacb53585db8d7ec2262e28dd2d8aad0c2ac27ef0388b0c50c2c6ec41

Observation d93ccb00-eaa5-42b8-b393-a84488cab151 · inbound

Subliminal Learning is Non-Semantic Distillation cites this paper.

Subliminal Learning is Non-Semantic Distillation Learning by Distilling Context

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:24:40.201293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:24:40.201293Z digest=sha256:41745b134fa7ef6f64fa6c32e2510d8743494c1a7410878c11c136f6cdc0c645