Pith. sign in

Paper Citation Record · LEDGER

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents

As of 11 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2608.07068.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07068 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:27:03.159262Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy32
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 455b6738-6ef9-474b-99e9-ea36e1c183e9 · outbound

This paper cites 2024 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , journal =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.546815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.546815Z digest=sha256:b476bdebf8f87e3c665aa46e9ea1aad2200a9231e32646344ac82cbb46dd8323

Observation 0abb4361-efe1-46cf-b5ad-0552bd94c96e · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.721471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.564209Z digest=sha256:fd0ce3c70357ddbc330c7c32e35d93c0f81a3b789bb7d970b29cff3dded97f13

Observation f13d4a9b-23a3-4578-bde9-c09b85266190 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.695715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.576330Z digest=sha256:6f36eb83ccbc1e1abe4a8961e790a8cc21af3177ea483cd6aa63165f7f803ebd

Observation a4505395-1095-4e43-acbe-d2b012487241 · outbound

This paper cites 2022 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , booktitle =

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.673866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.582963Z digest=sha256:b59ae57c1ddd0993a59d7197fd21e97224fe095bb6f38c3f7487194169dbc110

Observation 70cff33b-14d8-4a04-a565-0c1248209580 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.657063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.590920Z digest=sha256:17a7b09c8f1d2717ce6eb5ce6a14eb395c8fd847aead79f91d1f4e284b18e193

Observation 1f5e1797-024c-4d65-806e-a966deb5e865 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.639211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.597356Z digest=sha256:6e2852748ce52618e9f7840eea2a55ebf88bbe145bb0260af1c3ddc67f26e1ec

Observation 45ebc4e1-7248-4698-8403-d8287b7ed10b · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.618077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.605196Z digest=sha256:544ee57f9848f7d132841883101831c2e513fbff2da39b7ff64fa2d23ce1ffa5

Observation 873391c3-a8f0-47ae-bb56-761351afe39a · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.596596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.613927Z digest=sha256:f3f298448399ecc8a7699f7e216a0659a1a7bd2626f292c8abcd62eded82dc8a

Observation 96982bcc-1f0c-42cb-861f-303e7e8a7dbe · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.566135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.763900Z digest=sha256:14f5a3a2ed6fa817c288fd142b1e0f207e66ea860a1209b66bbbde80e9a722f3

Observation 2b7513d8-ba27-4f9b-ac18-db2de7bc204f · outbound

This paper cites 2022 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , eprint =

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.546896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.769634Z digest=sha256:22349fae6cbf7193b59cc13a35f74266258f66f993fdf9ea20167c9d98665c8c

Observation a928de82-b576-4505-97ef-8b8065eb9bb9 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.525647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.775875Z digest=sha256:f01d393c419f38fa6366e4de481d24405cbe8f34d397e884e5d8c7f0eebd5aae

Observation 0c9707db-d0e4-4d69-9465-a9dd4268dbbc · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.506651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.782245Z digest=sha256:84326487e2767d0665a44ba127916400bb736b1201c73a8b45b7ae293f167d51

Observation 480125f4-3b24-4135-a0cd-188be5e978fd · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.788263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.788263Z digest=sha256:5b91eab7609c604870d17579f00233dc3f4fd7cd9a5a59dd1161a78007d84870

Observation b518b6be-4309-4428-84ee-2786a09ec79e · outbound

This paper cites 2017 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2017 , booktitle =

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.484658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.795561Z digest=sha256:f5089e497dc477808b82d3ccdba8508de11b3332b4a600e78e4ea673b44ba6b2

Observation 857485bd-edc9-4c52-9aba-fd86e8bc2f1d · outbound

This paper cites 2020 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , eprint =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.802209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.802209Z digest=sha256:3c5bd59a45e388478f5e50e23d49662f298dc64993de22bcdba46693abd7d984

Observation 0b068964-4aaf-4c8a-974a-eaf0870c3a2b · outbound

This paper cites 2024 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , journal =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.808845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.808845Z digest=sha256:97c6937bc00e1a305e9b87df342e78dd2f0f35c5ba94e458d26a423c17c3cb52

Observation 853eca14-689b-45c3-bd20-626cb0263686 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.452615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.817512Z digest=sha256:c3cd55c0f13c68048457c3e13107a062a157ba2be16b3139b0d0bffb92a353e9

Observation cdeef7c9-36ce-4c34-895f-7505e8053fdc · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.435032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.823129Z digest=sha256:276f2a84f11daead52b48ac8f76ffcdeda51895221c12b61d73efa4c22fa075d

Observation 1d719ff2-dd7c-4048-a9e3-41e5444348c4 · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.415755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.829114Z digest=sha256:f1b697067fcb851d34ac2ee7b0fcba7c1c91f2ee4dd76b9fc1c53a356bcc014d

Observation b38e8396-3387-4c4e-bebb-b7e31d1a51f6 · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.391589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.836339Z digest=sha256:199eb5e024b4fd7b32fa886e37f00b57a02bb617e3237f3f3ab2e119901c4f8e

Observation 0a67c8b8-6960-4da5-a3d9-9566a850a489 · outbound

This paper cites 2020 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , booktitle =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.365844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.843933Z digest=sha256:ca289c93f8ae5178a9b7abc1307da8a6e9b5807fdf7eeae44fa2471a997d8e06

Observation 4d576344-e73a-428b-8311-0928715da227 · outbound

This paper cites 2020 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , booktitle =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.339213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.849621Z digest=sha256:07d7bdd7e8e84ad17f06aa6bf4f69aeceacb11e20264bb34c100e39124976677

Observation c70c19bd-c958-4a3d-bb73-95db1a076111 · outbound

This paper cites 2022 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , booktitle =

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.320789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.855671Z digest=sha256:4ab0c4ca4ab7bd71994700f8cff8da474e216c28dabd510d2240bf33fbeee647

Observation e7c85042-abb7-4347-aad2-3bfbda51bae5 · outbound

This paper cites MemoryBank: Enhancing Large Language Models with Long-Term Memory.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemoryBank: Enhancing Large Language Models with Long-Term Memory

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.862201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.862201Z digest=sha256:0817406d389fa4b0ecc3060a09939f376687a6813c10caa1cdc655aa936f91ad

Observation 1a887771-c3a8-4672-a380-dae2b0924863 · outbound

This paper cites MemGPT: Towards LLMs as Operating Systems.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemGPT: Towards LLMs as Operating Systems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.869014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.869014Z digest=sha256:f4686c0f34223e64ba796417c2b320f1f8070fd108b1fb4463a04f0f21d03a58

Observation 4c5247c1-f557-45b6-b6d1-54a814de5c68 · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.877464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.877464Z digest=sha256:78eb6b27452cc2e2e3fed088634b6202e72cf65e22b927082c4ca4802c8e18aa

Observation 2d6e29ba-9c35-4fd1-972b-8f954497babd · outbound

This paper cites A-MEM: Agentic Memory for LLM Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents A-MEM: Agentic Memory for LLM Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.890482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.890482Z digest=sha256:1ebf3f3dbd6a4ac7bfacb6dbb95fead6485f897b5971fb64a5907157dfda085d

Observation d14e6074-3da5-4849-9c36-11d616848838 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.298895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.898573Z digest=sha256:594098c81b842d5fd80e28c1c0bc7e5fba7f82a8bd08a3f164cb239366e58988

Observation a46602d0-e1bf-4530-8a47-8b83e5c88dc8 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.272600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.909411Z digest=sha256:3c0432b8be1b391c65210dfd24991adf2bf79cd7466efbbfd625172d193bd21a

Observation ea23424f-25b2-4855-a42d-17ab88a1917b · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.914748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.914748Z digest=sha256:c0d0d0dbc5add99626bf92258479cec6e9fe546bcf585eef1733654e82fcc177

Observation 93e39aca-0f0e-4f35-ae72-7bcf3e0f8eee · outbound

This paper cites MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.922284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.922284Z digest=sha256:cc931aa2fc494513022b0988f7f5f1f879e97ba56c1574f57611bf042a447138

Observation d39107a6-a2c9-4b40-869c-7c6d1040f5dc · outbound

This paper cites 2015 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2015 , eprint =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.929206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.929206Z digest=sha256:f06dd1dafa008675567ab38efc9e1407114720e5d1cb8ce188699343de070339

Observation 7961b108-5526-4762-abf4-973144734f3b · outbound

This paper cites On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.936313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.936313Z digest=sha256:b7fc3dca0de09826327210723f3451dcf0259fdab1f9c36fb2e2a7c4c1a58ba3

Observation 0abb995a-a53b-4c31-a726-08241a72f75e · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.942832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.942832Z digest=sha256:934a25e26e6ee00ff5485d0a971379d76c717991a1fb8931c071fb43c3092682

Observation f7fafce9-d68c-41bd-8011-a0f365dfb80b · outbound

This paper cites 2017 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2017 , eprint =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.949037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.949037Z digest=sha256:345bcf5f59ba9ba93e92ceae6e2d5c92842796eedc1b7dbc8fcead72de71b8dc

Observation 5a2f15c4-f24e-48d9-9c12-b43636033891 · outbound

This paper cites 2018 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2018 , booktitle =

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.229904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.955004Z digest=sha256:470db34c3b1e8e7c3307fc0b426806f663e52a4637297bd284230cb2b6bbcb59

Observation 37694dd9-00ff-4e41-82ca-ca53058e2bd2 · outbound

This paper cites 2019 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2019 , journal =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.962552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.962552Z digest=sha256:8587e3b27155dbaa407567e2e62b0285e7bd26570166c967caa75f226dbbbb27

Observation bb7f2561-e1ed-4b51-acc5-9b54b19f8ec3 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.972443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.972443Z digest=sha256:a82a3535c1320b27d37a9722252d587fe34188748e37d1e024f3190d7c41e863

Observation 852b807b-9e4f-42ac-bf7f-8c1b656fdc9c · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.201409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.977888Z digest=sha256:1c4610061e66a5afe9f5aef159e44c3e95e864260fb1ff2bdded05a7918d16b5

Observation 2024b10b-d0ce-4980-a4f8-0d6ff705e6fe · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.182215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.984336Z digest=sha256:0e1a99b69a95abba4c4d12d6e1d92572acf4ce44c06cae2b6135fa8f7724149b

Observation 78e84e00-9b3f-4942-958b-6693ce9e91ff · outbound

This paper cites 2024 , url =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , url =

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.156793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.991736Z digest=sha256:6ae7bcc48dde5d42aa1343e2deaad325e86f890730cfb8589edef48a7ed118ea

Observation 30ec3b9d-8b77-41e0-a369-d6f6f9d200a6 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.997944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.997944Z digest=sha256:ca742023264aee0ed9a45706e668d6660695aef355e9de4f2e6bbd29ceb05319

Observation 7a4e9157-d672-4fe1-97e4-c67db8f6b14d · outbound

This paper cites 2023 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , eprint =

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.138951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.003722Z digest=sha256:1bf83444a67adff103498fb192fb7599b83559fb9351bf6bd91685a0ddc4daba

Observation 6ba9ce27-d14c-4411-8882-1257b6b7f43e · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.122059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.009418Z digest=sha256:afdd24b2f2448c4aba53bb1ef093255b0e47259cca67619b7bcdac4398cf3480

Observation 5fd7106e-ae55-4f4e-8712-3326ccfcdc2d · outbound

This paper cites Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.015120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.015120Z digest=sha256:6c378154a281781eae7b5b032b6aa833303117aa43330c7f4416239da3d8e259

Observation d05900b5-00ec-485b-81f5-76ca848bdda9 · outbound

This paper cites 2025 , doi =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , doi =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.023759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.023759Z digest=sha256:293071dc9c7b53cf837d868bcbcbe745614a768e1e025d04c6dd296d6d4e9e86

Observation 4423f137-4b36-44b7-b5de-72c7a235c443 · outbound

This paper cites GAIA: a benchmark for General AI Assistants.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents GAIA: a benchmark for General AI Assistants

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.028808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.028808Z digest=sha256:e046766ced3aa88d8d61387b78ba328bc122ab4508bb0fb0c7465cf5cd7071ec

Observation 900a15df-f44f-4d1a-8f55-4a490c4e7c91 · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.034622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.034622Z digest=sha256:4c2f0c6d75e3c53c8ee6e01a599cb0b6835c5473434571d08bf9f98b9b2032fd

Observation 0ab5ee6d-5006-4113-bf0c-2aafd94fb84c · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.040495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.040495Z digest=sha256:271d7dc302e3057158c0a22a288cdc7776158f39e80a4c16fa8ff73ba91719ae

Observation 3e890958-464a-46cb-9b64-129dd7c46e51 · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.048699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.048699Z digest=sha256:66bb6d7312f303d13d84fc4840da302e0da6b5f2d697fc2fbae4fcdee11fd148

Observation e145e79c-8be9-4d5f-a57d-21e8f70caf6f · outbound

This paper cites MemPO: Self-Memory Policy Optimization for Long-Horizon Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemPO: Self-Memory Policy Optimization for Long-Horizon Agents

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.055159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.055159Z digest=sha256:385f3e9e76212deabd9a7b3363e564b055126976c4e8ee8becdbb749456109eb

Observation 5ea944fa-e5cc-419e-8d30-92ccfce8dd37 · outbound

This paper cites Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing , pages =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing , pages =

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.062670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.062670Z digest=sha256:4fb3e3d5e1a0e1685c3483a3f86fe8bd461ae795609e0c682b0795e05589c070

Observation 1eed30bd-e77a-4bbf-8542-4c0ff148063a · outbound

This paper cites 2024 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , eprint =

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.092543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.071247Z digest=sha256:b6aab46ed5180c6a858ead513504efed089d6cc0f06e909fee9c15c608efebc9

Observation 63bf59ab-0664-4efc-8a5f-bbfae2fbc33c · outbound

This paper cites Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , series =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , series =

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.077783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.077783Z digest=sha256:e8c8c391b88f5714f41033525afe39f27454911ad58457620038217563940a84

Observation f592d38c-cecd-4fe3-80e4-942c4e0b34d8 · outbound

This paper cites DistiLLM: Towards Streamlined Distillation for Large Language Models.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents DistiLLM: Towards Streamlined Distillation for Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.084232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.084232Z digest=sha256:ddf430fb38ae31bcd81d9624da8ad1fbebe98e6b51512d5a9327057b1fc987a6

Observation a18067dd-93e8-4003-8b89-213948e5fe54 · outbound

This paper cites Knowledge Distillation: A Survey.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Knowledge Distillation: A Survey

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.092957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.092957Z digest=sha256:2be544f72ac63d16ae9aa3b786d15f6307a7356ac3bdb2280d4083482e516207

Observation c1dc625f-2db1-4bd7-a19b-e1f63b1338c8 · outbound

This paper cites Training language models to follow instructions with human feedback.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Training language models to follow instructions with human feedback

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.100686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.100686Z digest=sha256:8f8a3f62b494042430d396cdbdef94eaba2f396526d47f18b8e225d2f1e02eca

Observation 4f51b412-d0bf-4259-beb0-81892f28fada · outbound

This paper cites Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology , year =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology , year =

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.107712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.107712Z digest=sha256:6b581515772cd3bd6f453ac6951c3187f95c9c9d2f49724e24c62679156c270a

Observation 5d8922ee-6806-48eb-878d-89440a30131f · outbound

This paper cites A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.113238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.113238Z digest=sha256:934f411ac73e5c24a1a74a6a76222087b4578385b7c5d2dd755bb3d2538ea2e8

Observation f2c13c4a-5ffe-4121-8948-f0309e808f58 · outbound

This paper cites LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.120485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.120485Z digest=sha256:52d5ed36adfff94890a23c42d7df661b6552bbdd87515762d44d6ee47a36569c

Observation 5d4a1d03-f9e2-47a7-8a75-925617dc56fb · outbound

This paper cites LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.125713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.125713Z digest=sha256:9aa6a96f5be71f900fe840780d64e5ba2adfb2815876819f98f15b000343f7cf

Observation 246598c8-e0f9-4b28-ab4b-1fc376aba6b7 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.058462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.130931Z digest=sha256:01bf06a8d3e26f9ab9fc461c21b85a629456657f22bb00c97518ed48c6eaaa18

Observation aa13d68d-873f-4589-a4fc-43b2910deec4 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.040052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.135799Z digest=sha256:ea5a212c6973d14cba1593372b3b7c6f5a9960dc2d5d7a45d5451a777a825a0e

Observation a3c21e61-b154-4505-a0d7-9f88c0316599 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.021336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.141522Z digest=sha256:805d421591dc56dd9b661e6f44fea6d1ad67c04af2d494e21c6292bb282304ca

Observation 9ce9fc88-cc2d-4f85-bbb7-20e01e3547a3 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.003901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.147576Z digest=sha256:b5759d2a875ed9614a6eedf65cea31aa348c14a465df71381fabf25c436f03ab

Observation 916ce6ef-50c4-4289-8e1e-4cc2303c0ea9 · outbound

This paper cites 2025 , eprint=.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.159262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.159262Z digest=sha256:d307766bf12960e0ef7cff520eac3b36335bd2c8b2a423312f6ec7c1cf297be8

Pith citing papers

No inbound Pith citation observations are available.