Pith. sign in

Paper Citation Record · LEDGER

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents

As of 11 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2608.07068.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07068 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:27:03.159262Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy32
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 455b6738-6ef9-474b-99e9-ea36e1c183e9 · outbound

This paper cites 2024 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , journal =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.546815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.546815Z digest=sha256:96a45de829a68b47a75fdbd7b97bd40b8172e7644179ee6262a8b6d44950e9d4

Observation 0abb4361-efe1-46cf-b5ad-0552bd94c96e · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.721471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.564209Z digest=sha256:f0f02af13202a24e2537acbeec6bf3e58606d7f1e9f072fa42d0927235519b62

Observation f13d4a9b-23a3-4578-bde9-c09b85266190 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.695715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.576330Z digest=sha256:dda9f76892ebe366485727271513a3455fb0654a27b854612fad6a94865e7924

Observation a4505395-1095-4e43-acbe-d2b012487241 · outbound

This paper cites 2022 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , booktitle =

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.673866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.582963Z digest=sha256:75f033c74f1f4261da0415efcbb2ec0fd1afd38c57ad13ae8fd13ccd0a8bf6ab

Observation 70cff33b-14d8-4a04-a565-0c1248209580 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.657063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.590920Z digest=sha256:e839ce83bb46fd0455eaacd83474c3d84c4d47ceda623686eff4612d47936cfd

Observation 1f5e1797-024c-4d65-806e-a966deb5e865 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.639211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.597356Z digest=sha256:6abc11521e3d45e735cece36e6a8ab95a5f8114a818ab98d6358129f935f5b35

Observation 45ebc4e1-7248-4698-8403-d8287b7ed10b · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.618077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.605196Z digest=sha256:c2a5fa95f421ab476182345198d775c37e06a2d85bfb6c8012db721351def0f4

Observation 873391c3-a8f0-47ae-bb56-761351afe39a · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.596596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.613927Z digest=sha256:629ffc655e49d5e33c05cb3ba0291c0e4f94780d63d6615698087d126816e7ca

Observation 96982bcc-1f0c-42cb-861f-303e7e8a7dbe · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.566135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.763900Z digest=sha256:659e9ae6d074378e18cba3d695c769407116bf5f21ce6569d6720912ebdff13f

Observation 2b7513d8-ba27-4f9b-ac18-db2de7bc204f · outbound

This paper cites 2022 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , eprint =

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.546896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.769634Z digest=sha256:298dfda580468a884bf75611a5f5e9671a21f579c817518871df7f1910710ae9

Observation a928de82-b576-4505-97ef-8b8065eb9bb9 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.525647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.775875Z digest=sha256:6610ed939e850afa6ac54ca0dd537b11337f0c18fb461aca7b17e3ab35d281db

Observation 0c9707db-d0e4-4d69-9465-a9dd4268dbbc · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.506651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.782245Z digest=sha256:57c21431507a960f0c6b5baabced20ac87a0acf19cc9ad7dbea00487d33784fe

Observation 480125f4-3b24-4135-a0cd-188be5e978fd · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.788263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.788263Z digest=sha256:c7526a9140721030918172f1353fbb09aae25452b063e2dc26aecf9f014b5b72

Observation b518b6be-4309-4428-84ee-2786a09ec79e · outbound

This paper cites 2017 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2017 , booktitle =

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.484658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.795561Z digest=sha256:de75ea91d33ebc2f7ff38297c08cdc248da6db8aa51fbee606de033077ee0f47

Observation 857485bd-edc9-4c52-9aba-fd86e8bc2f1d · outbound

This paper cites 2020 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , eprint =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.802209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.802209Z digest=sha256:3a6537aee1125592c2a8c489b1661ca48588e08e907da66be4181d82151034b3

Observation 0b068964-4aaf-4c8a-974a-eaf0870c3a2b · outbound

This paper cites 2024 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , journal =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.808845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.808845Z digest=sha256:dbe7f871d5da11c23ec477fdbfc4be25bce676121a0e9beb9e2e6fc45bf70d54

Observation 853eca14-689b-45c3-bd20-626cb0263686 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.452615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.817512Z digest=sha256:73f7aa9995aa0ed2d577bef1bafa7291658c1737fc794e7d321a934f0ac704a5

Observation cdeef7c9-36ce-4c34-895f-7505e8053fdc · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.435032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.823129Z digest=sha256:a8eccdbeab83d494c98f78475998241d2d7347c0df4ca0ba53607438d9b1aa27

Observation 1d719ff2-dd7c-4048-a9e3-41e5444348c4 · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.415755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.829114Z digest=sha256:4d1bfaa945f6540742f50e8da15ca27f6c40e086578c1572a07a62cc0178f82e

Observation b38e8396-3387-4c4e-bebb-b7e31d1a51f6 · outbound

This paper cites 2025 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , booktitle =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.391589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.836339Z digest=sha256:b0daa4f5211b1a4798d33ba5f3ea1504c75583696eefcf799cc66fa38ac70d2a

Observation 0a67c8b8-6960-4da5-a3d9-9566a850a489 · outbound

This paper cites 2020 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , booktitle =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.365844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.843933Z digest=sha256:a7276c662cae4e1dd457db385fa1069ee9ebdb7ab81e68596fe20b042f334d5c

Observation 4d576344-e73a-428b-8311-0928715da227 · outbound

This paper cites 2020 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2020 , booktitle =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.339213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.849621Z digest=sha256:0e373afbe5120ff3382909668b1ed3c921e090735bc2bf1b8dc5754d94e58728

Observation c70c19bd-c958-4a3d-bb73-95db1a076111 · outbound

This paper cites 2022 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2022 , booktitle =

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.320789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.855671Z digest=sha256:d38e507cd891e60f597da30dc989d05e61907ec2fbb02ce98a2e9fedc67ec04e

Observation e7c85042-abb7-4347-aad2-3bfbda51bae5 · outbound

This paper cites MemoryBank: Enhancing Large Language Models with Long-Term Memory.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemoryBank: Enhancing Large Language Models with Long-Term Memory

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.862201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.862201Z digest=sha256:1a11780f45786b7c2a2d7c4c819e8b5c7240548d094746a027b21e5cd825ef37

Observation 1a887771-c3a8-4672-a380-dae2b0924863 · outbound

This paper cites MemGPT: Towards LLMs as Operating Systems.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemGPT: Towards LLMs as Operating Systems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.869014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.869014Z digest=sha256:0d59fd2ca008a63e332e086aa64fb7f41cec809b9990a3b40f57039db40c371c

Observation 4c5247c1-f557-45b6-b6d1-54a814de5c68 · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.877464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.877464Z digest=sha256:cb2408de7a3bc3c367420523adf67e1d7b4d29938116220f58915cbe8318db60

Observation 2d6e29ba-9c35-4fd1-972b-8f954497babd · outbound

This paper cites A-MEM: Agentic Memory for LLM Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents A-MEM: Agentic Memory for LLM Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.890482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.890482Z digest=sha256:fce744df5d4879b1c4439f775cb4d9cc6ff34e5d00fb911949de9c828b4472f6

Observation d14e6074-3da5-4849-9c36-11d616848838 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.298895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.898573Z digest=sha256:8bffc9d19d8127bfd1d8624878f82a7d1bc96ad0da6e9e0e7224295e763c8df1

Observation a46602d0-e1bf-4530-8a47-8b83e5c88dc8 · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.272600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.909411Z digest=sha256:b992be712ae23363576bb2f01b906822aac8115badf348c21296eb4c0e7803a4

Observation ea23424f-25b2-4855-a42d-17ab88a1917b · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.914748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.914748Z digest=sha256:9482dcb8f4b6d68fe187301c4b3b10d1a59db86b9629762563c6379f90e13475

Observation 93e39aca-0f0e-4f35-ae72-7bcf3e0f8eee · outbound

This paper cites MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.922284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.922284Z digest=sha256:c97fdd174c4598c06bb81d4683b2071e65db36cb835506c7308daba7ad116604

Observation d39107a6-a2c9-4b40-869c-7c6d1040f5dc · outbound

This paper cites 2015 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2015 , eprint =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.929206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.929206Z digest=sha256:c311eeb87c6938763d71cf18132becd164cce203699cefa940d2b7409f7444e9

Observation 7961b108-5526-4762-abf4-973144734f3b · outbound

This paper cites On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.936313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.936313Z digest=sha256:7d2f98f7031d5e45b116c1406323caacd4f29c34475540c3541c6a21627b0c3a

Observation 0abb995a-a53b-4c31-a726-08241a72f75e · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.942832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.942832Z digest=sha256:9be4d63ad245ad71b19673a44b67c60dafe3525e77bc88903a7c055f6d88a016

Observation f7fafce9-d68c-41bd-8011-a0f365dfb80b · outbound

This paper cites 2017 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2017 , eprint =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.949037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.949037Z digest=sha256:b3d59c2d4402f6f34f8d17a05088c5f44209c55b5e5d05724442fb9f2ce56f0b

Observation 5a2f15c4-f24e-48d9-9c12-b43636033891 · outbound

This paper cites 2018 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2018 , booktitle =

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.229904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.955004Z digest=sha256:198c586edcc2c6573905fd8552ca0df6307c99871fab18e2dad2c9df29fc98cc

Observation 37694dd9-00ff-4e41-82ca-ca53058e2bd2 · outbound

This paper cites 2019 , journal =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2019 , journal =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.962552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.962552Z digest=sha256:9470c1e7dc8a63ef580107524f90ac9d2380dd56c57a1ea4a0259de2f820d759

Observation bb7f2561-e1ed-4b51-acc5-9b54b19f8ec3 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.972443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.972443Z digest=sha256:e2b9a7e73b699d330ed1cc97c50a6e5d325c658632e56380d79485707537c80c

Observation 852b807b-9e4f-42ac-bf7f-8c1b656fdc9c · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.201409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.977888Z digest=sha256:10b472d3f805d19eda4a5595d4877ffc631b443f4fcd16ce736c476bf20a4f4c

Observation 2024b10b-d0ce-4980-a4f8-0d6ff705e6fe · outbound

This paper cites 2023 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , booktitle =

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.182215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.984336Z digest=sha256:111d6d6172b1b920ead50ea34f2935b00e43cee1f246cc415cd7ca2042885fc1

Observation 78e84e00-9b3f-4942-958b-6693ce9e91ff · outbound

This paper cites 2024 , url =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , url =

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.156793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:02.991736Z digest=sha256:6a441bd9b3003f587a9e2e44c68f5ad0ae9c0c700727f0ce6dfac80535563110

Observation 30ec3b9d-8b77-41e0-a369-d6f6f9d200a6 · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:02.997944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:02.997944Z digest=sha256:83920fa2ed28cd69fa956c9140a4bd7ec64e3ecd250c0473ecbb6feb6cde3e3c

Observation 7a4e9157-d672-4fe1-97e4-c67db8f6b14d · outbound

This paper cites 2023 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2023 , eprint =

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.138951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.003722Z digest=sha256:5379b6b6f4a218835b2f3a5dbae5bb7976f8ec8192dd1462a299629f700f13ea

Observation 6ba9ce27-d14c-4411-8882-1257b6b7f43e · outbound

This paper cites 2024 , booktitle =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , booktitle =

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.122059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.009418Z digest=sha256:81c64b1cc9d24ac2c7febe2afb58a44d122786651cefe78684faeda77b222f4e

Observation 5fd7106e-ae55-4f4e-8712-3326ccfcdc2d · outbound

This paper cites Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.015120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.015120Z digest=sha256:2fd8cc39e277b4e70ce8be6a30cd3c45c482d60c2dd6ebd98e3a7f90a580847d

Observation d05900b5-00ec-485b-81f5-76ca848bdda9 · outbound

This paper cites 2025 , doi =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , doi =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.023759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.023759Z digest=sha256:d254b5511df474b0dc298b071b2225b72e073f4de09433203eca71260720494d

Observation 4423f137-4b36-44b7-b5de-72c7a235c443 · outbound

This paper cites GAIA: a benchmark for General AI Assistants.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents GAIA: a benchmark for General AI Assistants

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.028808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.028808Z digest=sha256:4d87c2b6b4b672ba1c1bafc67daae868cae8951758c785f7dae142bdfb36f0db

Observation 900a15df-f44f-4d1a-8f55-4a490c4e7c91 · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.034622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.034622Z digest=sha256:cde22a4e1a4a555219125baf1a811c8872cf5864ac2bca36ce060c39544616ef

Observation 0ab5ee6d-5006-4113-bf0c-2aafd94fb84c · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.040495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.040495Z digest=sha256:2788f3b8ff2a3980bba0eb82693213ec2b269b782319076912a95188e8480274

Observation 3e890958-464a-46cb-9b64-129dd7c46e51 · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.048699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.048699Z digest=sha256:fde771aa9d601c7227fc279d1e4e6ddb5231532807715d73d2ecbbaf346d5c99

Observation e145e79c-8be9-4d5f-a57d-21e8f70caf6f · outbound

This paper cites MemPO: Self-Memory Policy Optimization for Long-Horizon Agents.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents MemPO: Self-Memory Policy Optimization for Long-Horizon Agents

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.055159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.055159Z digest=sha256:ca081b801ece8909b7eb57bb67bfc61e606139fac4af721c5b3e3453c13e9e54

Observation 5ea944fa-e5cc-419e-8d30-92ccfce8dd37 · outbound

This paper cites Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing , pages =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing , pages =

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.062670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.062670Z digest=sha256:3eb1a699bb1d3b9dcdc04ee6c5df376abc2a79e849486e910f73425cce6db4b6

Observation 1eed30bd-e77a-4bbf-8542-4c0ff148063a · outbound

This paper cites 2024 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2024 , eprint =

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.092543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.071247Z digest=sha256:4469a6273fb32f262f1c8d8444a4206c52deda7ef3f0dac7502f7ba9dd0dfb91

Observation 63bf59ab-0664-4efc-8a5f-bbfae2fbc33c · outbound

This paper cites Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , series =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , series =

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.077783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.077783Z digest=sha256:cc2e3b8cfd8a7d6b229900ace3bc40316ad05a5515967012ba88a6f3a74fd9e3

Observation f592d38c-cecd-4fe3-80e4-942c4e0b34d8 · outbound

This paper cites DistiLLM: Towards Streamlined Distillation for Large Language Models.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents DistiLLM: Towards Streamlined Distillation for Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.084232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.084232Z digest=sha256:b9aff4630313499fe500cf48160cdc58322bb239ff2e7b55ff25a5b38afa0f00

Observation a18067dd-93e8-4003-8b89-213948e5fe54 · outbound

This paper cites Knowledge Distillation: A Survey.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Knowledge Distillation: A Survey

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.092957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.092957Z digest=sha256:f0217c3e4b622ba259128027e1d4e027ab7fec85f5fa5c794686d834518c159b

Observation c1dc625f-2db1-4bd7-a19b-e1f63b1338c8 · outbound

This paper cites Training language models to follow instructions with human feedback.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Training language models to follow instructions with human feedback

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.100686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.100686Z digest=sha256:22a275ef590ab6030dd003fcc4629eb1d0ee23c8d816c2611da297c8eb04bd2c

Observation 4f51b412-d0bf-4259-beb0-81892f28fada · outbound

This paper cites Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology , year =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology , year =

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.107712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.107712Z digest=sha256:2b8831ece5b7ba43176693be50923e3588b65bd5df37b15d3a8ef5d507c50816

Observation 5d8922ee-6806-48eb-878d-89440a30131f · outbound

This paper cites A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.113238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.113238Z digest=sha256:6091d9cf5b713684a6da755e715a365740eeee36f2c3f5597339e04516a14e57

Observation f2c13c4a-5ffe-4121-8948-f0309e808f58 · outbound

This paper cites LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.120485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.120485Z digest=sha256:ac2a78ecf7671eeabbda467e79d17e4824397ffe64339f74c3fa62a9f270eede

Observation 5d4a1d03-f9e2-47a7-8a75-925617dc56fb · outbound

This paper cites LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.125713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.125713Z digest=sha256:ad52d13b96028e2fb87611e0be18549dd32a09446c6aad09ff30fba1a5a846dc

Observation 246598c8-e0f9-4b28-ab4b-1fc376aba6b7 · outbound

This paper cites 2025 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint =

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.058462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.130931Z digest=sha256:1d7172e83b406f60f0b8e45ae9b2b9611c922b6c853957cb6e578dbab01c58dd

Observation aa13d68d-873f-4589-a4fc-43b2910deec4 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.040052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.135799Z digest=sha256:e1f90c91c56901ef2eb79ca8f1980139eb26e1991df4d171ba2f5c0994fbde30

Observation a3c21e61-b154-4505-a0d7-9f88c0316599 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.021336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.141522Z digest=sha256:08b08ef5853cdbd7ad916a4505501232aa0b46c56b8af0fb77149e4a69bbb901

Observation 9ce9fc88-cc2d-4f85-bbb7-20e01e3547a3 · outbound

This paper cites 2026 , eprint =.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2026 , eprint =

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:27:04.003901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T15:27:03.147576Z digest=sha256:59a5c77f42ade845b813c1d7d08e124d6c5ac1899c2b20b8564d8e611c31a576

Observation 916ce6ef-50c4-4289-8e1e-4cc2303c0ea9 · outbound

This paper cites 2025 , eprint=.

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents 2025 , eprint=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T15:27:03.159262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:27:03.159262Z digest=sha256:670d90afe1fa003991763b321c7560e7199c59cef6fe842c3deb673cdd924d43

Pith citing papers

No inbound Pith citation observations are available.