Pith. sign in

Paper Citation Record · LEDGER

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.01475.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.01475 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:47:51.893563Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved49
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a36002a7-5bac-42bd-affb-97ac55fbee3c · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.991948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:00.510118Z digest=sha256:71b1ceae1c0a3773a7c9e4dc7b637868c39fe24e0b9b2338dc3665194475a9ee

Observation 622a48e0-de72-472a-a0c6-f43cbd5a0d5a · outbound

This paper cites FireAct: Toward Language Agent Fine-tuning.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.567143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.567143Z digest=sha256:bd9c35c46a5589d9835a5b7e2da86c64d7873428c83fba802f90f86aa6ffd9da

Observation cdb255ec-c282-429d-b500-d0974b2c4670 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.920232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:00.600539Z digest=sha256:8b4dd2e31dfd5335faf7c9dba951e1995d245bfa4e6116bdf034deb3dafc8b46

Observation 09c23844-b8de-457a-8f49-b541a72ad153 · outbound

This paper cites The Llama 3 Herd of Models.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.634668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.634668Z digest=sha256:ed0d80667d25cbb4160b723e4a5edcdeef58f81a11f22854a689512be17ebbd1

Observation c9a98375-1970-4930-a02b-426f3af5ff02 · outbound

This paper cites AgentRefine: Enhancing Agent Generalization through Refinement Tuning.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.667874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.667874Z digest=sha256:73c6f26d6e81fa0ca5d2d3f4d9f96aa144358f02c067591c79c3bc91557a4060

Observation d1554dac-de26-4c17-a6d1-f9e56b6d969e · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.702676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.702676Z digest=sha256:7c43deb2055b2f76a96a78a986d0a9a717e2b158027535a46127c32b3cb7e944

Observation 4f1f6fff-9af9-4d3e-b4ac-53ce6e465792 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.747128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.747128Z digest=sha256:7d4607e840331d2ffd00d5d413dc8fca10fcbb33dd44a4e336360d118cac88bb

Observation 2d03f915-c74c-48f6-80cf-1086413519b4 · outbound

This paper cites Understanding the planning of LLM agents: A survey.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Understanding the planning of LLM agents: A survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.789154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.789154Z digest=sha256:681367f279dd06cc3b36450b6cd4783eed47fa07bb2b0bd393f6f0d045148096

Observation 145139b7-6a41-41b1-83c9-2d07506fb036 · outbound

This paper cites Mistral 7B.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Mistral 7B

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.031991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.031991Z digest=sha256:49a0ef77cdab24c3be182afe7fabbbeaba5a19149100bf29ba0c7dceea1c3e1b

Observation 991364c6-71ba-430b-9574-4ba73a98a6cc · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.594789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.594789Z digest=sha256:f084c7b2ab78eaa264f54628ac3de6008b3a095ab457ff5e0879bf3643ea8198

Observation c8cd9652-6ac1-47a8-b7eb-d497f3294b5b · outbound

This paper cites Formal-LLM: Integrating Formal Language and Natural Language for Controllable LLM-based Agents.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Formal-LLM: Integrating Formal Language and Natural Language for Controllable LLM-based Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.662536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.662536Z digest=sha256:37ce27effa01a43457b80c09b86d737fb09812ea4a92a1c624a75408d5c33d18

Observation bdf620be-78d7-460a-8495-b6706a267e70 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.863579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:49.716341Z digest=sha256:c8fae7a7219812b567ee5b23083c63eb261b24de1a3601f94b76ce28c1e8503a

Observation ba4523f0-6713-44a6-b0b7-af57f0cb6422 · outbound

This paper cites LLM+P: Empowering Large Language Models with Optimal Planning Proficiency.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization LLM+P: Empowering Large Language Models with Optimal Planning Proficiency

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.774743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.774743Z digest=sha256:b8917021059ac9ed342b1acf4e3e38bfc2ee8368c6db79b8c1c671b77e269257

Observation 5af744e2-fe28-4738-b2fc-489e74b80782 · outbound

This paper cites BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.833363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.833363Z digest=sha256:237010b8a79334a0a431c44afdf1ac252ed616b99aebf0f5f4a5f40c2300c356

Observation a65dd59a-a1b4-411f-b430-cb9d5ee88428 · outbound

This paper cites AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.917199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.917199Z digest=sha256:1b4ac8c1b04c473cb7ac4d9f99761135796dd3f6ea32e2042ea96cf77828f164

Observation 332c5315-02f6-4918-b037-9877450aa16c · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.843640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:49.986828Z digest=sha256:c6af9a67ae9f1f93c2d157ff968c1e1f20e10518dfa83c5f993bd477c578b442

Observation 2e3e4396-081c-46b7-96e1-774db6aa8875 · outbound

This paper cites Iterative Reasoning Preference Optimization.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Iterative Reasoning Preference Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.045904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.045904Z digest=sha256:dabaec672a4af6b1c7788b26539daf6f4ec56ddce822f199276497e7e0547298

Observation 4745af46-8475-4974-8e1f-1c4c094354d4 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.820253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.078733Z digest=sha256:a0a431e5ad7bd1a33a5292861312aef31131acac8edd9ac3304b4e55f0e53155

Observation 350e6912-6345-47ef-9f5d-6cac0f3a2371 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.793120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.148555Z digest=sha256:e10a5ea69cfadb9d0bd209a2c2fc3e95cc44e5382d05c1e2e916de3c3c6f2677

Observation 5079c6a6-cdb6-41fd-b1ab-9f9ae3cc8a4c · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.769821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.207706Z digest=sha256:0de711c8c8c236a2190efdfe6a9f024479bbca92e62cf821e00b6ac2c6d13e71

Observation 931bd661-e97d-4f6f-967b-f24bcb3b0f54 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.267644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.267644Z digest=sha256:0dbcbd04e75971f81b63ef587d2798670317cbc19045a9201b810e1a86743a72

Observation cea3d569-e049-4f60-a72b-a5d4fdeed13b · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.736404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.310634Z digest=sha256:2474a8c23e56477aa1d66898b2a5dc05918684049d02bdfe5cf1c7ffd0eec299

Observation 73379426-c257-4d6d-af78-5ded0d9fbf6f · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.708647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.357062Z digest=sha256:708deb34d86352feaea659e1ea44de78040ae4db4501239acc699e7439396904

Observation 4a9cee47-1aa2-4991-a7b1-7859d5b8e63e · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.676944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.408001Z digest=sha256:40e6024f46581c2cfc9298a7ec6677ac23a12cd2710d7e3273da6b0a53a2d2db

Observation 64644980-8ae8-475e-abf1-d0179d441fd0 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.456516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.456516Z digest=sha256:fbc7d826cadf133abe84d0f7e16b5756a329ab4ce222f84d6ce56f9bfda1b9d0

Observation 23f5e70d-6c64-403e-bb6b-de896eb297d7 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.629814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.492430Z digest=sha256:1471e052fdb1d9e1cc0fd5b012509ce8992c250f63af8ecb4ebd6a47855001b2

Observation 0fb23917-fac1-4484-b599-61a0918d3c2e · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.598303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.552291Z digest=sha256:ce6b85bb4ca79fe790c59b550b8a0ee3393ccb418e56fd7695c10d6e51c70cc4

Observation 16145647-c582-436b-b3d2-29220fe5fe15 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.545699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.586993Z digest=sha256:4cab39a7c6375de31c1d5ad74a2887a04dd16649ed46ae2ba1aa352d41164c9d

Observation 71e2d4e4-b78d-4a1c-a11b-653df120ae42 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.631039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.631039Z digest=sha256:f4ebd6350125334303b77f615e95968bc1f470a2a4cf2bc02e97b4b1601c7b38

Observation 2140e262-1074-482b-af60-d775f7fce3b1 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.517420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.667589Z digest=sha256:147bfc5d4bd11bbd3981058e5904d629ca6898852e1eac11fab1355a231f6062

Observation 2446a275-e63f-4c1e-a0a0-90e6b32ed609 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.482921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.718877Z digest=sha256:9086d2c5b1530b290a4d35c2c476ad72a85adeae3b5ec34822037e5fadbe433e

Observation 8e05df81-66a4-4122-be45-da0781c574ff · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.454199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.761020Z digest=sha256:bfba2f28b5378d3ec8dd60efd2f96ace49cb28c9033ad2beadefd9098738c6e5

Observation c4321bbf-c7f3-42b4-90a5-8aca749254dc · outbound

This paper cites Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.812929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.812929Z digest=sha256:03cce8253f679a68019b99d635bc06589b238172628e50fb5f57c76c02728547

Observation 850942b7-7b5d-49c8-bca6-142653ce5f88 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:50.865043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:50.865043Z digest=sha256:5c49b45effff6acfc0529973c24c24c20db6a605063fe57884637e838f5b0225

Observation 427674e7-135b-4c20-add5-a291ce001290 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.414023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.899805Z digest=sha256:dfcb03e873388530f95e61f2e350271136e47bd68a997448d275f1dc35d93762

Observation b920b238-47a8-41b8-a720-b299ae11feeb · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.386404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:50.938845Z digest=sha256:01b58e06d7850985c759e2b69ff46d7d46dad385ca576ed885e7121256e06b21

Observation d8b74d7d-38b9-4532-b0f1-3b92aad717fa · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.359706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:51.013837Z digest=sha256:3afd6857f8bd6460d7ba8b021ce95db531a9d531b2a516fd7b02211cbd0f9371

Observation 6144216c-00b3-4443-bd3b-686053f31473 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.088342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.088342Z digest=sha256:ce7a0390f634ab8e0ca40c9610c572c56b99d2ccabed89e1f9b241c0b840061c

Observation a684049a-f95e-4239-8b00-b51f5d2c69a8 · outbound

This paper cites AgentGym: Evolving Large Language Model-based Agents across Diverse Environments.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.170109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.170109Z digest=sha256:e1af6347291189dcb937aca19ffb4bd2e1072c8f85817909659d70c6e9f64c73

Observation 22913d76-7fc6-4849-b3bc-136441fc9911 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.327626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:51.261249Z digest=sha256:273b3675cc001fdbadb901c0741bc0ee551ffd0ade7a4e3343f3761727009fb2

Observation 7596d7b1-f3ef-428c-a474-82d1cdf0b5c4 · outbound

This paper cites Qwen2.5 Technical Report.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Qwen2.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.387390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.387390Z digest=sha256:98969130812f1b66f8a2e7a8e402529e5dcb341dabb0c51eead25d86ae386dc7

Observation 825e833e-7986-4d95-8019-6472786c8528 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:48:04.306526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:51.487455Z digest=sha256:25d2f932e8014842c9bc0331cc5644790047d8c6fc444dd7999fb6208bb58ed0

Observation 9ceab20c-1743-4f48-b571-657861aa7ba1 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.572633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.572633Z digest=sha256:31d53e0421fdb057282f1e7e4c672d12992b0901d8fe5b1e0e595389e481dbf2

Observation 04f4195d-9466-4ca0-941e-56761dc095d6 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.659712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.659712Z digest=sha256:aec83e9cf23e4ea81c49f6f25754abe4a092c27127450a514c2d13d8b165d617

Observation 69a5f091-aa01-45de-a69e-8fbe41d9fae7 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.766466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.766466Z digest=sha256:0fed6b98bc8151335fa89f7c4a179b477ceb5270c49d617b9329fbf791b548d5

Observation 370278cf-386c-497b-95ca-2d64b6804765 · outbound

This paper cites an unresolved cited work.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:47:52.230431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:47:51.798750Z digest=sha256:a211174b1bf5a52927d8da0b5d946f33fd958638a2f75ab6e887745b7f931267

Observation bf651055-7644-4e73-97af-9589fc132d70 · outbound

This paper cites LaMMA-P: Generalizable Multi-Agent Long-Horizon Task Allocation and Planning with LM-Driven PDDL Planner.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization LaMMA-P: Generalizable Multi-Agent Long-Horizon Task Allocation and Planning with LM-Driven PDDL Planner

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.821700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.821700Z digest=sha256:0ad8395373329bac6a07fcc4ad32b797b0d9d3d7c7fb867d7b7528c0ef822925

Observation 97feef45-44db-42a5-9660-e46edeba4a60 · outbound

This paper cites online" 'onlinestring :=.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization online" 'onlinestring :=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.859589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.859589Z digest=sha256:737d717d3fc2185db392d51c070ae6d956a5991f2d2ecb3da07af4a6b9ae28b3

Observation deed7873-ab9b-466f-b0bf-db31a38ac115 · outbound

This paper cites write newline.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization write newline

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:51.893563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:51.893563Z digest=sha256:93f1f5efc40ef66e8e12417084d09b1c3a52584bae37fea4002cfcf2bd4d506b

Pith citing papers

No inbound Pith citation observations are available.