Pith. sign in

Paper Citation Record · LEDGER

Reinforcing Language Agents via Policy Optimization with Action Decomposition

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2405.15821.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.15821 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:37:47.590688Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T01:48:28.550913Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8029fb17-4eaa-44da-a0b9-37a8b23ba8c9 · inbound

ProgRM: Build Better GUI Agents with Progress Rewards cites this paper.

ProgRM: Build Better GUI Agents with Progress Rewards Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:37:47.590688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:37:47.590688Z digest=sha256:38ccb217bb49fcc5e420f850e7a6a534d0dbc55ba32e388b6373114438569bee

Observation c6e8579a-c8a1-4ccf-b5e2-58311378a9dd · inbound

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models cites this paper.

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:30:58.541659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:09:36.341574Z digest=sha256:764611d25a5ed81ae1770e106c9b2a8efc50b7fc3dee7af5349f600ee6a089d2

Observation ee7f9348-e1a5-4dd5-b9ff-df2570a48912 · inbound

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy cites this paper.

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:48:28.553279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T01:46:24.724553Z digest=sha256:6e6b79b342de9b00c156d848da990860bccd686a9ea3d655ba502cded2053b96

Observation 053a9567-4e62-4d7c-8e96-0d99d0544034 · inbound

BiCAA: Bidirectional Credit Assignment for Search-Augmented Agent cites this paper.

BiCAA: Bidirectional Credit Assignment for Search-Augmented Agent Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T00:23:00.148968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:23:00.148968Z digest=sha256:be2df05695a0cf4041f3b4368e24ec714b612c075c042cca1af4b13301ef98d8