Pith. sign in

Paper Citation Record · LEDGER

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents

As of 11 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2608.07068.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07068 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:27:03.159262Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy32
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 455b6738-6ef9-474b-99e9-ea36e1c183e9 · outbound

This paper cites 2024 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , journal =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.546815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.546815Z digest=sha256:98cfa696f051254eba4a083c38b8001f6041374076d89b0997b22448ef92760f

Observation 0abb4361-efe1-46cf-b5ad-0552bd94c96e · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.721471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.564209Z digest=sha256:aa241bf532a5c108683fa1e7205790067f896c7e128af53a8ff2eae7159f6e46

Observation f13d4a9b-23a3-4578-bde9-c09b85266190 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.695715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.576330Z digest=sha256:91affe7a7d0d6b958f6b9f8e3cf3004b042e0dec6d9056447163bada0f2e929a

Observation a4505395-1095-4e43-acbe-d2b012487241 · outbound

This paper cites 2022 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , booktitle =

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.673866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.582963Z digest=sha256:7e981909ff13f4777188ebef8cc853f1b4b9931eeab0cb66b3831cb4ca7a6d73

Observation 70cff33b-14d8-4a04-a565-0c1248209580 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.657063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.590920Z digest=sha256:5f905758c424a268ffce5ff406eddb37a3c8afb151a4a6a7c0ddc3222d703558

Observation 1f5e1797-024c-4d65-806e-a966deb5e865 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.639211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.597356Z digest=sha256:00d65643f4b0eb80c1f4e9c3fdde047e0e34a250f4b494212ca6ff95e4944b09

Observation 45ebc4e1-7248-4698-8403-d8287b7ed10b · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.618077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.605196Z digest=sha256:d41fec31166342a7e1ae538b5019efdb5578966c4448d5ea9ff49a733e8dc3b2

Observation 873391c3-a8f0-47ae-bb56-761351afe39a · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.596596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.613927Z digest=sha256:eceaaf42baf0ec98f9e0ccbf60ffec523366ed6f64cc1208b7870193ef86db3b

Observation 96982bcc-1f0c-42cb-861f-303e7e8a7dbe · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.566135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.763900Z digest=sha256:abff0933482ce812253bc13b959709327ae156d9776f8055538bfef6bedc1618

Observation 2b7513d8-ba27-4f9b-ac18-db2de7bc204f · outbound

This paper cites 2022 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , eprint =

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.546896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.769634Z digest=sha256:0ed47abdc1f74d40517de6d102f0a4cd1bf964942a450ecf98657b605047a64f

Observation a928de82-b576-4505-97ef-8b8065eb9bb9 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.525647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.775875Z digest=sha256:bc9357b7f8bb0e6b86ee01b75bde6421c085fdff2a59c8d12283086edc927e7d

Observation 0c9707db-d0e4-4d69-9465-a9dd4268dbbc · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.506651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.782245Z digest=sha256:b6554c44fc017666e3b80fbde4208fe880125ac31850921f8e743d419aa8a5cc

Observation 480125f4-3b24-4135-a0cd-188be5e978fd · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.788263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.788263Z digest=sha256:e616db735177e5cbcb9049505f985323e7c198c9273a55eaf5f03ce4ec14dbb2

Observation b518b6be-4309-4428-84ee-2786a09ec79e · outbound

This paper cites 2017 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2017 , booktitle =

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.484658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.795561Z digest=sha256:b7edb87032cfe6a3b85cafdee5bfc757bd293f17dfd47cd0d1151453e01500f1

Observation 857485bd-edc9-4c52-9aba-fd86e8bc2f1d · outbound

This paper cites 2020 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , eprint =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.802209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.802209Z digest=sha256:89210f736f4256003830f73e28cb4898f3d1f905435f69f1c4d47c853a5a7590

Observation 0b068964-4aaf-4c8a-974a-eaf0870c3a2b · outbound

This paper cites 2024 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , journal =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.808845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.808845Z digest=sha256:5ded7de180d983e8797f79f171fd99fedf2bebe68846b6b0fa9d31745cf24436

Observation 853eca14-689b-45c3-bd20-626cb0263686 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.452615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.817512Z digest=sha256:999ec981fa0f4ab793abe62ee48782faace85ad79a5fa6e0046919d1d5149ac3

Observation cdeef7c9-36ce-4c34-895f-7505e8053fdc · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.435032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.823129Z digest=sha256:6e1080ee907916c8f85c3acb84c9b3ee23b42546f82aba540ed6f609e4c78720

Observation 1d719ff2-dd7c-4048-a9e3-41e5444348c4 · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.415755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.829114Z digest=sha256:40f15ba439abc65474d5eaf44d36903dd73cad73e29dede41f6044f125a2459c

Observation b38e8396-3387-4c4e-bebb-b7e31d1a51f6 · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.391589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.836339Z digest=sha256:19355855bd5ae5ab70ccb6f1be8b5269a3766311271a95717b49b6acbec855dc

Observation 0a67c8b8-6960-4da5-a3d9-9566a850a489 · outbound

This paper cites 2020 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , booktitle =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.365844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.843933Z digest=sha256:b770a27a73fa06202f6664c3b0708d266e8fd0dbdac12eda5465fa1cf24833a1

Observation 4d576344-e73a-428b-8311-0928715da227 · outbound

This paper cites 2020 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , booktitle =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.339213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.849621Z digest=sha256:8037f7ce66398c94e5629f133dbebe877c4eb2d4dd28a0534221429461733fa2

Observation c70c19bd-c958-4a3d-bb73-95db1a076111 · outbound

This paper cites 2022 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , booktitle =

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.320789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.855671Z digest=sha256:2edb9f6f0159aae3293f58db8a9ee1eeafcd5d06a7157724ed6aa5cc969a1b82

Observation e7c85042-abb7-4347-aad2-3bfbda51bae5 · outbound

This paper cites MemoryBank: Enhancing Large Language Models with Long-Term Memory.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemoryBank: Enhancing Large Language Models with Long-Term Memory

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.862201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.862201Z digest=sha256:c83190864b7c15335c5cc021ccd763b7d3447f27d43d6bd94d7231fd9c8b2fdb

Observation 1a887771-c3a8-4672-a380-dae2b0924863 · outbound

This paper cites MemGPT: Towards LLMs as Operating Systems.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemGPT: Towards LLMs as Operating Systems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.869014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.869014Z digest=sha256:5015f011360fa73d4f979a3c8741876650461af718935568ff04a6f8191f10dd

Observation 4c5247c1-f557-45b6-b6d1-54a814de5c68 · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.877464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.877464Z digest=sha256:3e4500ae206c555c68d03fd71837d825780ab5b8b2de6e7716ae39b80e570b86

Observation 2d6e29ba-9c35-4fd1-972b-8f954497babd · outbound

This paper cites A-MEM: Agentic Memory for LLM Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents A-MEM: Agentic Memory for LLM Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.890482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.890482Z digest=sha256:487c1bd2736f475275e4577edf92c9c6c7fb2363f3cda3e198278ee2d10de7ee

Observation d14e6074-3da5-4849-9c36-11d616848838 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.298895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.898573Z digest=sha256:20fffe798a269dd092c96fc40d533f7a19afbc390d022510950057213fb411ef

Observation a46602d0-e1bf-4530-8a47-8b83e5c88dc8 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.272600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.909411Z digest=sha256:152c75d9ad7fe7c5782ec95cc6d0f6ee0e27eccdf7ddf701bf383cb5756054fe

Observation ea23424f-25b2-4855-a42d-17ab88a1917b · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.914748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.914748Z digest=sha256:c10115cafd073f84c1c878b188817dbd2c79762e8f02d6ab310d3ff2f5e502dd

Observation 93e39aca-0f0e-4f35-ae72-7bcf3e0f8eee · outbound

This paper cites MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.922284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.922284Z digest=sha256:e3d0486741845b8695a294c9b4bf05dcd8aeefb8862855aff7b683a43163fd5f

Observation d39107a6-a2c9-4b40-869c-7c6d1040f5dc · outbound

This paper cites 2015 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2015 , eprint =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.929206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.929206Z digest=sha256:dee42c554f7fc6dc6199e930948ce8ba0da182bbb18d880a6652362066375a7b

Observation 7961b108-5526-4762-abf4-973144734f3b · outbound

This paper cites On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.936313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.936313Z digest=sha256:fe930b87f931efba455c89cc79bfa2c25ee1c6dbc5c940d27f652a4badfccd1b

Observation 0abb995a-a53b-4c31-a726-08241a72f75e · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.942832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.942832Z digest=sha256:9f7fba870035130f52388bb0f5f8c1a65d8bd62a20683c4b25a75a4de072fa3b

Observation f7fafce9-d68c-41bd-8011-a0f365dfb80b · outbound

This paper cites 2017 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2017 , eprint =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.949037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.949037Z digest=sha256:fee3ca5490b52d8c583d0cfa22a3917b991fd034c5a4edf8478f43fcdcead08d

Observation 5a2f15c4-f24e-48d9-9c12-b43636033891 · outbound

This paper cites 2018 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2018 , booktitle =

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.229904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.955004Z digest=sha256:5753b9970f92206a5c1c419693d08349d8e071bb41a211bd4176b29e7f3d08ba

Observation 37694dd9-00ff-4e41-82ca-ca53058e2bd2 · outbound

This paper cites 2019 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2019 , journal =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.962552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.962552Z digest=sha256:76290aeab2832e3b0aa8a7ac71765407accc65174f5c6c15e80f6dbcb7a8650a

Observation bb7f2561-e1ed-4b51-acc5-9b54b19f8ec3 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.972443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.972443Z digest=sha256:03ef4ef3690f6bee4f17ead46f71f39471a7f5ede99a245b5db974e9b230db76

Observation 852b807b-9e4f-42ac-bf7f-8c1b656fdc9c · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.201409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.977888Z digest=sha256:5093594cfb2ea6bd6187d989db761278965206fe9832122481b358e6a48d3172

Observation 2024b10b-d0ce-4980-a4f8-0d6ff705e6fe · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.182215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.984336Z digest=sha256:36e90747276d0503fd57693758c985b462323c0b8ce412927f34beb5f9202b92

Observation 78e84e00-9b3f-4942-958b-6693ce9e91ff · outbound

This paper cites 2024 , url =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , url =

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.156793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.991736Z digest=sha256:18f7e93082cc00f00f7168f15157ef41021b39f145ad4f40661602ef6077cfef

Observation 30ec3b9d-8b77-41e0-a369-d6f6f9d200a6 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.997944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.997944Z digest=sha256:ed9520bf14325f7898d83ca81ee9b9952aee737b06b265b04b595ab3dbd945d3

Observation 7a4e9157-d672-4fe1-97e4-c67db8f6b14d · outbound

This paper cites 2023 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , eprint =

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.138951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.003722Z digest=sha256:bfe26b756eeb4ee63c42b54afc6bf7fe9cfad5db52331a1e573022302233808f

Observation 6ba9ce27-d14c-4411-8882-1257b6b7f43e · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.122059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.009418Z digest=sha256:b3def4ae386f5ea47f606f938b13daeb3e09705d03e27c815e60a0deab0302d9

Observation 5fd7106e-ae55-4f4e-8712-3326ccfcdc2d · outbound

This paper cites Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.015120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.015120Z digest=sha256:1ec7efa3273f91d9fa21d7fb92d9a19d9f2d5b4e20d3920b924fac504edcec3e

Observation d05900b5-00ec-485b-81f5-76ca848bdda9 · outbound

This paper cites 2025 , doi =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , doi =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.023759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.023759Z digest=sha256:6f4c4cee8a7a324fcac9cf5c33fa3e71e072c17d49ce903da7a2cf7c59a54bb6

Observation 4423f137-4b36-44b7-b5de-72c7a235c443 · outbound

This paper cites GAIA: a benchmark for General AI Assistants.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents GAIA: a benchmark for General AI Assistants

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.028808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.028808Z digest=sha256:bb9a4604462b8fb82a116fc0f37fc2f7d3f39e905deefe6e9632922789c7c3ca

Observation 900a15df-f44f-4d1a-8f55-4a490c4e7c91 · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.034622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.034622Z digest=sha256:8c57a3c84e585f89e9857d1dc7ac3b2aa7503e5d19a3006c516827233e265ff2

Observation 0ab5ee6d-5006-4113-bf0c-2aafd94fb84c · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.040495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.040495Z digest=sha256:96800995faf87a2b2d4ba82af1134c7e0bb6edd50293265abde4ff5d6e942a52

Observation 3e890958-464a-46cb-9b64-129dd7c46e51 · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.048699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.048699Z digest=sha256:fbe3489558d808e3bdb41c5c7394d84bbd044e77f3099bd0f99cd98715090997

Observation e145e79c-8be9-4d5f-a57d-21e8f70caf6f · outbound

This paper cites MemPO: Self-Memory Policy Optimization for Long-Horizon Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemPO: Self-Memory Policy Optimization for Long-Horizon Agents

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.055159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.055159Z digest=sha256:32687d31c4814fc4515f3d43421c9b8fc0e0f9db30951796b1a992984e5f1fd0

Observation 5ea944fa-e5cc-419e-8d30-92ccfce8dd37 · outbound

This paper cites Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing , pages =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing , pages =

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.062670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.062670Z digest=sha256:985e02ea06a2dc241493df5402cc36ceb742f48b4bd1445dcf8d7623a9120519

Observation 1eed30bd-e77a-4bbf-8542-4c0ff148063a · outbound

This paper cites 2024 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , eprint =

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.092543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.071247Z digest=sha256:d396599b474487f945f44188e018dfee769637b87fe4fbe47b0e56d001f8aea4

Observation 63bf59ab-0664-4efc-8a5f-bbfae2fbc33c · outbound

This paper cites Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , series =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , series =

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.077783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.077783Z digest=sha256:8eb9334ebd36a8112edcda84725cfe061c033e0bd6911aa4f8707f2734c90b43

Observation f592d38c-cecd-4fe3-80e4-942c4e0b34d8 · outbound

This paper cites DistiLLM: Towards Streamlined Distillation for Large Language Models.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents DistiLLM: Towards Streamlined Distillation for Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.084232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.084232Z digest=sha256:99d6d17edde6810650114710ae81a6a58af9ae3c93a935ee8ec2f45a51d47196

Observation a18067dd-93e8-4003-8b89-213948e5fe54 · outbound

This paper cites Knowledge Distillation: A Survey.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Knowledge Distillation: A Survey

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.092957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.092957Z digest=sha256:95bdb80273ded01cdd75a41a87cb5cb3dcbc19e31f368c7c103604699cf69fff

Observation c1dc625f-2db1-4bd7-a19b-e1f63b1338c8 · outbound

This paper cites Training language models to follow instructions with human feedback.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Training language models to follow instructions with human feedback

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.100686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.100686Z digest=sha256:94228553660dbd2be23e67d6605c9a2236b611b897671e0e58a04e23b805033c

Observation 4f51b412-d0bf-4259-beb0-81892f28fada · outbound

This paper cites Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology , year =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology , year =

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.107712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.107712Z digest=sha256:871294e9c49ed4be3374f36415268809383b0e5b6e3ee86b90c2e33bdccb3e06

Observation 5d8922ee-6806-48eb-878d-89440a30131f · outbound

This paper cites A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.113238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.113238Z digest=sha256:ad588eb7c0146a50fe681bf08ceb16785964a145415269cd602875cd98fbd3de

Observation f2c13c4a-5ffe-4121-8948-f0309e808f58 · outbound

This paper cites LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.120485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.120485Z digest=sha256:02ccfdc3200e70745a3b31f7f6a993d120087f1cc71b7755cb129564d2bbfdaa

Observation 5d4a1d03-f9e2-47a7-8a75-925617dc56fb · outbound

This paper cites LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.125713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.125713Z digest=sha256:8a26d8c775948dc12198d4dc3d328ff7998c690a1f562ce5c06f7612f3b5fac3

Observation 246598c8-e0f9-4b28-ab4b-1fc376aba6b7 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.058462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.130931Z digest=sha256:9663ac9e59e6d7e01d7c05d40de1d35869a4ccd2156eb001aa5595cf6a557dcd

Observation aa13d68d-873f-4589-a4fc-43b2910deec4 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.040052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.135799Z digest=sha256:c93e805bd837aa5e03a1af474cc9aa20eb61c1d030013bb3e80905ed7dc60942

Observation a3c21e61-b154-4505-a0d7-9f88c0316599 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.021336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.141522Z digest=sha256:495be79b845c319da47263c192890036430342e98efb1985717353c8bab26e2f

Observation 9ce9fc88-cc2d-4f85-bbb7-20e01e3547a3 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.003901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.147576Z digest=sha256:45321137e077e8452cd6acaf82de5503449e0cdf708b7f6ee567dca32c3274a0

Observation 916ce6ef-50c4-4289-8e1e-4cc2303c0ea9 · outbound

This paper cites 2025 , eprint=.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.159262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.159262Z digest=sha256:9de80b82eff672ba55a82316ef0aaef44bb76e8ade037f0dac61484bd6fb4971

Pith citing papers

No inbound Pith citation observations are available.