Pith. sign in

Paper Citation Record · LEDGER

Reinforcing Language Agents via Policy Optimization with Action Decomposition

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2405.15821.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.15821 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:26:14.134503Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T01:48:28.550913Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1366ad71-e52c-4167-afaf-22dcf5d9518d · inbound

Multi-Agent System for Cosmological Parameter Analysis cites this paper.

Multi-Agent System for Cosmological Parameter Analysis Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T05:26:14.134503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:26:14.134503Z digest=sha256:32ac28131f81d766813d71ae933dc905beafffb6a212b9c95cd3d2070a97d3a5

Observation ab31dc5f-bfb7-4906-b197-5504d5cb11da · inbound

Aviary: training language agents on challenging scientific tasks cites this paper.

Aviary: training language agents on challenging scientific tasks Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T23:07:33.621669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:07:33.621669Z digest=sha256:2f480d5ba91d43a743c5a55a72c16bd418a34a0e65ecfba992a7c3ab5525b40f

Observation 8029fb17-4eaa-44da-a0b9-37a8b23ba8c9 · inbound

ProgRM: Build Better GUI Agents with Progress Rewards cites this paper.

ProgRM: Build Better GUI Agents with Progress Rewards Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:37:47.590688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:37:47.590688Z digest=sha256:1f2de8e327fa6ce614bd400f0e112bb8c6f0674e5cd9b8dc873e822669c94724

Observation c6e8579a-c8a1-4ccf-b5e2-58311378a9dd · inbound

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models cites this paper.

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:30:58.541659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T17:09:36.341574Z digest=sha256:becd7923cbadc3ec1aab47794139be0c03a333eaafd19d4d7779cc449dc091ce

Observation ee7f9348-e1a5-4dd5-b9ff-df2570a48912 · inbound

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy cites this paper.

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:48:28.553279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T01:46:24.724553Z digest=sha256:1245796854e763bbbdce646fec8e285592664c4de74ae12a0f7d027c9f8585a4

Observation 053a9567-4e62-4d7c-8e96-0d99d0544034 · inbound

BiCAA: Bidirectional Credit Assignment for Search-Augmented Agent cites this paper.

BiCAA: Bidirectional Credit Assignment for Search-Augmented Agent Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T00:23:00.148968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:23:00.148968Z digest=sha256:632009527bb2ec690a4e803ebe3e45b199d83f059c547c7a293a488023b16760