Pith. sign in

Paper Citation Record · LEDGER

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents

As of 11 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2608.07068.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07068 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:27:03.159262Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy32
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 455b6738-6ef9-474b-99e9-ea36e1c183e9 · outbound

This paper cites 2024 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , journal =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.546815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.546815Z digest=sha256:f1e053e7531b8c11dbc8ac9f39a9f9307f5a88219bcb841a095633488ceb1028

Observation 0abb4361-efe1-46cf-b5ad-0552bd94c96e · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.721471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.564209Z digest=sha256:281ad66f7dd15a1737d55905aa3da703cfa8052c3c55cd68a3d57b48e608717d

Observation f13d4a9b-23a3-4578-bde9-c09b85266190 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.695715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.576330Z digest=sha256:21a48fb4ff33c1497b311b0ab921698cce539d6b5a2abc70edf93e0c370c99ff

Observation a4505395-1095-4e43-acbe-d2b012487241 · outbound

This paper cites 2022 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , booktitle =

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.673866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.582963Z digest=sha256:4fa5e8c7b693e84b290394380ef41ce1d48e2a78c8c41e37f1aafda3f079dd96

Observation 70cff33b-14d8-4a04-a565-0c1248209580 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.657063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.590920Z digest=sha256:33d311502c9cab2a58587cfa051ed4913396c9705de895813f80b30a19c145ad

Observation 1f5e1797-024c-4d65-806e-a966deb5e865 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.639211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.597356Z digest=sha256:6ebd2afc03dc64826e0e2494a089c3519b5acee49ab5f9cc32ed18f63fbcf3cf

Observation 45ebc4e1-7248-4698-8403-d8287b7ed10b · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.618077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.605196Z digest=sha256:7d6c1da3ff714691a45412d4c1d8f59c9423c8966ef8ac89d48124dbd5e8a5b9

Observation 873391c3-a8f0-47ae-bb56-761351afe39a · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.596596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.613927Z digest=sha256:464b6e31b86f717f575c3adcc8b307a79d0d906de7f5d3a6d8c59e35e901ada4

Observation 96982bcc-1f0c-42cb-861f-303e7e8a7dbe · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.566135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.763900Z digest=sha256:a87cabc68beaf04977f7c1b873f55d4dcee4225fa38c4504e9b04d2f2d50f596

Observation 2b7513d8-ba27-4f9b-ac18-db2de7bc204f · outbound

This paper cites 2022 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , eprint =

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.546896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.769634Z digest=sha256:ef2077ce2eb55f906531ed683df858191d15987e018a7f5e2e633fcb13cebb2c

Observation a928de82-b576-4505-97ef-8b8065eb9bb9 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.525647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.775875Z digest=sha256:5234ba493e0fa934ee5f0e7173ebcbf9f0b7b8bdd82af0363961efa987569950

Observation 0c9707db-d0e4-4d69-9465-a9dd4268dbbc · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.506651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.782245Z digest=sha256:36a6ca550692ece2e2f39f0992d0f9e2c6bce0d1010c6acc268315ea8b2935f2

Observation 480125f4-3b24-4135-a0cd-188be5e978fd · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.788263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.788263Z digest=sha256:f425f07a3c94d59469224ecedebab84b7a26de80fb07e7227115feaf87018dba

Observation b518b6be-4309-4428-84ee-2786a09ec79e · outbound

This paper cites 2017 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2017 , booktitle =

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.484658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.795561Z digest=sha256:c86e76c5c309202ecbdedae13f0c0dc2f69f0fc244251a1b11f77d7282944f84

Observation 857485bd-edc9-4c52-9aba-fd86e8bc2f1d · outbound

This paper cites 2020 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , eprint =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.802209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.802209Z digest=sha256:341ab5da3a454189acdded68046ea8722187fbb3c2ac3723e6c1e7417791649b

Observation 0b068964-4aaf-4c8a-974a-eaf0870c3a2b · outbound

This paper cites 2024 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , journal =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.808845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.808845Z digest=sha256:eb5aaabef7102adf5993a27370885b654269e1fece97d87b83cfa018fd9bf378

Observation 853eca14-689b-45c3-bd20-626cb0263686 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.452615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.817512Z digest=sha256:2c50461336929282d480ae13b3c5140a569909a625858b2599f0be94bd55b489

Observation cdeef7c9-36ce-4c34-895f-7505e8053fdc · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.435032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.823129Z digest=sha256:7b0d1838b95fb60ee958abc9a48bbc3efa8d277260672679de3138b90d9908f0

Observation 1d719ff2-dd7c-4048-a9e3-41e5444348c4 · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.415755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.829114Z digest=sha256:2c6501b1b75936036e44fecfc401f5dc298c899f7aa4ace40ef9a6be952ee912

Observation b38e8396-3387-4c4e-bebb-b7e31d1a51f6 · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.391589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.836339Z digest=sha256:2feb0acdce6338fa2d3fd76c6dcec3f8380502629146797ff93b6b48b992cf4d

Observation 0a67c8b8-6960-4da5-a3d9-9566a850a489 · outbound

This paper cites 2020 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , booktitle =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.365844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.843933Z digest=sha256:91b4f8887857d77ca53f4af2cb2434eddfd25a064635b742cd6c0058703546e3

Observation 4d576344-e73a-428b-8311-0928715da227 · outbound

This paper cites 2020 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , booktitle =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.339213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.849621Z digest=sha256:923e49bb0846acc95d0cf6f7a58f8806ae43d17cfe23cc15224cbe150dca9b4c

Observation c70c19bd-c958-4a3d-bb73-95db1a076111 · outbound

This paper cites 2022 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , booktitle =

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.320789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.855671Z digest=sha256:d7476ef24709262cfc2a072c42d09cf4236bc35fca4fca1178837d39c14c075f

Observation e7c85042-abb7-4347-aad2-3bfbda51bae5 · outbound

This paper cites MemoryBank: Enhancing Large Language Models with Long-Term Memory.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemoryBank: Enhancing Large Language Models with Long-Term Memory

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.862201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.862201Z digest=sha256:ca72e0aade2f0e4f4afa04a9921bf1385d4c9fc0712877b6dfce7c19f1715600

Observation 1a887771-c3a8-4672-a380-dae2b0924863 · outbound

This paper cites MemGPT: Towards LLMs as Operating Systems.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemGPT: Towards LLMs as Operating Systems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.869014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.869014Z digest=sha256:6c7271f18208261045dc88d5284acaa821b8f348c4cc6f9e24e75b4d7d3aeefa

Observation 4c5247c1-f557-45b6-b6d1-54a814de5c68 · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.877464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.877464Z digest=sha256:82e62a0f20ec17c5bb2f202c432ecdd1fc94080ad85e7f6581c367316f62e5fa

Observation 2d6e29ba-9c35-4fd1-972b-8f954497babd · outbound

This paper cites A-MEM: Agentic Memory for LLM Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents A-MEM: Agentic Memory for LLM Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.890482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.890482Z digest=sha256:429b9c57ee007b27289b7e11739aa688b3d2f3db54123c45faf21e84c7cfeb8a

Observation d14e6074-3da5-4849-9c36-11d616848838 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.298895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.898573Z digest=sha256:7ae7d75e57a512b07cbe2e827cef446a502edb301de5305805e5db73f6eb1d68

Observation a46602d0-e1bf-4530-8a47-8b83e5c88dc8 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.272600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.909411Z digest=sha256:2563aa97c14715a7993ae4520759cc1a06c830582f017692042c8357fca9771b

Observation ea23424f-25b2-4855-a42d-17ab88a1917b · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.914748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.914748Z digest=sha256:a897ff9473ddf5fc91b406c8cbc5365b96a4973dac4dc5a39c6a664cd022a7f3

Observation 93e39aca-0f0e-4f35-ae72-7bcf3e0f8eee · outbound

This paper cites MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.922284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.922284Z digest=sha256:8dcbc59553fe55a1847e59514a3ae65bb3d260e8090ee4190a9aee1a187d5a58

Observation d39107a6-a2c9-4b40-869c-7c6d1040f5dc · outbound

This paper cites 2015 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2015 , eprint =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.929206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.929206Z digest=sha256:ed1c5fedcc66ed78ab7412e78556d1bbcccf69bf854ec78364fd8807c8b3300c

Observation 7961b108-5526-4762-abf4-973144734f3b · outbound

This paper cites On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.936313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.936313Z digest=sha256:26a6921bd5f8d13d9cd2093a0578d16ef67a5cfc4f9e4dc6880f7fc5c3d3a7ee

Observation 0abb995a-a53b-4c31-a726-08241a72f75e · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.942832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.942832Z digest=sha256:f062cdd350deebc073e424a52425e842f2d216982df29d20774ef4a7dc7a5471

Observation f7fafce9-d68c-41bd-8011-a0f365dfb80b · outbound

This paper cites 2017 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2017 , eprint =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.949037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.949037Z digest=sha256:f8536a4baaf3de1e9240c3bf96066916fe286eaa5b615a36675f67e81b509bff

Observation 5a2f15c4-f24e-48d9-9c12-b43636033891 · outbound

This paper cites 2018 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2018 , booktitle =

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.229904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.955004Z digest=sha256:968b5c49cbdc09ee4f21c855139980d3d0d4ae2ce4a269810c303dece3232311

Observation 37694dd9-00ff-4e41-82ca-ca53058e2bd2 · outbound

This paper cites 2019 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2019 , journal =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.962552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.962552Z digest=sha256:be8040ea098e1e99d4943368799d2532706b190f9b49b5b631799d823a800bea

Observation bb7f2561-e1ed-4b51-acc5-9b54b19f8ec3 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.972443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.972443Z digest=sha256:fd32340f67774fa6a7cf344bb488d1872fb765e961b18f16c42563809eb9a001

Observation 852b807b-9e4f-42ac-bf7f-8c1b656fdc9c · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.201409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.977888Z digest=sha256:10b95f26a6ba6b50a726c9f528e90a52a7dba57c96cb304624b8f36de42d9cd6

Observation 2024b10b-d0ce-4980-a4f8-0d6ff705e6fe · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.182215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.984336Z digest=sha256:1e76b38e4ffb6e280a10b22ccc25bb093177d79cf5f2e65eef4b4928dd984808

Observation 78e84e00-9b3f-4942-958b-6693ce9e91ff · outbound

This paper cites 2024 , url =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , url =

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.156793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.991736Z digest=sha256:117bad1b6e65721c63a7b72e50f1e6d6f848270c06a4bbcbf208a6e711210f10

Observation 30ec3b9d-8b77-41e0-a369-d6f6f9d200a6 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.997944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.997944Z digest=sha256:543b6e6ab90b23d3f78c5a7a298e07780d72cea0fafcaac296d13650c58271a2

Observation 7a4e9157-d672-4fe1-97e4-c67db8f6b14d · outbound

This paper cites 2023 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , eprint =

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.138951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.003722Z digest=sha256:88a1a15266dd5c60ebb5b87ff433ed903f1fb438837bc9041d88a617a0901c29

Observation 6ba9ce27-d14c-4411-8882-1257b6b7f43e · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.122059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.009418Z digest=sha256:68dd41842f6bbc248b45b6a32b139ae9e37640bbd48301fe2e14e953334eed37

Observation 5fd7106e-ae55-4f4e-8712-3326ccfcdc2d · outbound

This paper cites Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.015120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.015120Z digest=sha256:9c18289d3652ddfe73807068e76ef0ba70937a05be8dc59505997a1df3f399b7

Observation d05900b5-00ec-485b-81f5-76ca848bdda9 · outbound

This paper cites 2025 , doi =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , doi =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.023759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.023759Z digest=sha256:b6dc853cd0c367d7a202489bd7714c1075797bcdecbccc29f33a7c5588699c9e

Observation 4423f137-4b36-44b7-b5de-72c7a235c443 · outbound

This paper cites GAIA: a benchmark for General AI Assistants.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents GAIA: a benchmark for General AI Assistants

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.028808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.028808Z digest=sha256:38092f43e5f832255ae9c79a54259a3fad2110ff9a26d6fbfbf54125830f7bbe

Observation 900a15df-f44f-4d1a-8f55-4a490c4e7c91 · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.034622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.034622Z digest=sha256:32fb0a2d6bb95646448c506e39cb3dff42cb0f4c7baf3dccd889fb3e20f77698

Observation 0ab5ee6d-5006-4113-bf0c-2aafd94fb84c · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.040495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.040495Z digest=sha256:dab215deb4d0276001534d1d2d940b96f6b1e880203b0d40333ecc72d8e56708

Observation 3e890958-464a-46cb-9b64-129dd7c46e51 · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.048699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.048699Z digest=sha256:31f049d5eb2b194df8663120bf9cdd4c2a3ed69637593cfaba0aa50f6526ac2c

Observation e145e79c-8be9-4d5f-a57d-21e8f70caf6f · outbound

This paper cites MemPO: Self-Memory Policy Optimization for Long-Horizon Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemPO: Self-Memory Policy Optimization for Long-Horizon Agents

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.055159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.055159Z digest=sha256:a84d8511d4bbc76b8a978f92a0415a715dba341213fb26b52b9f41d7c1d8d568

Observation 5ea944fa-e5cc-419e-8d30-92ccfce8dd37 · outbound

This paper cites Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing , pages =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing , pages =

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.062670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.062670Z digest=sha256:fe5f28453912f8fa4b74e2a594dedc21929bfc3442be4bc86645d574f722e340

Observation 1eed30bd-e77a-4bbf-8542-4c0ff148063a · outbound

This paper cites 2024 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , eprint =

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.092543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.071247Z digest=sha256:10c8752c31fd40a2f78452aaf4a5ebe8bf08e4b1ef311157454c1806cb787854

Observation 63bf59ab-0664-4efc-8a5f-bbfae2fbc33c · outbound

This paper cites Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , series =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , series =

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.077783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.077783Z digest=sha256:00e325528e92c5e450b63d80d049498cc74238a3eae853913c7a9e847941d8d6

Observation f592d38c-cecd-4fe3-80e4-942c4e0b34d8 · outbound

This paper cites DistiLLM: Towards Streamlined Distillation for Large Language Models.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents DistiLLM: Towards Streamlined Distillation for Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.084232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.084232Z digest=sha256:b3ef38ceb9b5d56a20931775a69232bbea89a5562c7d5b12615325005a6fe667

Observation a18067dd-93e8-4003-8b89-213948e5fe54 · outbound

This paper cites Knowledge Distillation: A Survey.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Knowledge Distillation: A Survey

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.092957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.092957Z digest=sha256:a2290b222f6c5afb5f60b3a48c7c1cf3da1b406c0cf94e2ecfd7a051ed532ad5

Observation c1dc625f-2db1-4bd7-a19b-e1f63b1338c8 · outbound

This paper cites Training language models to follow instructions with human feedback.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Training language models to follow instructions with human feedback

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.100686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.100686Z digest=sha256:19e8428b906f4005a8077db3f854e3a81131d423563e5f7cacaf0557b04e9a33

Observation 4f51b412-d0bf-4259-beb0-81892f28fada · outbound

This paper cites Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology , year =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology , year =

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.107712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.107712Z digest=sha256:f2d6f91e6164af7963646b64ac2bad99e48bc836e11c19dbb8c1dcddab9d6a56

Observation 5d8922ee-6806-48eb-878d-89440a30131f · outbound

This paper cites A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.113238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.113238Z digest=sha256:122f749f274bdf281938a5439b500d0cdaa4036fced7e1cb82b15667f25a4758

Observation f2c13c4a-5ffe-4121-8948-f0309e808f58 · outbound

This paper cites LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.120485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.120485Z digest=sha256:3ab4fb96a91f30821c5110a4dd7a4d00fb311930f030fd781750f32c14b8b2eb

Observation 5d4a1d03-f9e2-47a7-8a75-925617dc56fb · outbound

This paper cites LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.125713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.125713Z digest=sha256:f398cdfd018486617bdc843b7879a3c70c9d272cf3b6b3646a75a2e0d139e919

Observation 246598c8-e0f9-4b28-ab4b-1fc376aba6b7 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.058462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.130931Z digest=sha256:80e9959240250a38e687185308e5babd855a28dc246929ea845f71e324acc573

Observation aa13d68d-873f-4589-a4fc-43b2910deec4 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.040052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.135799Z digest=sha256:cc868600654a985f2a00812a1771412142e22c5b79fe86e850bd2d8252020721

Observation a3c21e61-b154-4505-a0d7-9f88c0316599 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.021336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.141522Z digest=sha256:52efd19dd8858333f683dbdb3de3b44b20710ed176246106ea6b71f9e4479c76

Observation 9ce9fc88-cc2d-4f85-bbb7-20e01e3547a3 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.003901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.147576Z digest=sha256:59f18bf46bc0df76a414fa808a2444db162d0a9c32adc40acd2a4c0d639b0337

Observation 916ce6ef-50c4-4289-8e1e-4cc2303c0ea9 · outbound

This paper cites 2025 , eprint=.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.159262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.159262Z digest=sha256:7e83da96f6b40d28fba87ab2df9b367af414153298c5d99498d0d37e2057746e

Pith citing papers

No inbound Pith citation observations are available.