Pith. sign in

Paper Citation Record · LEDGER

BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2308.05960.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.05960 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:24:04.475783Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

9
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bc17063a-9eec-4624-88dd-71fd05981727 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 170

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:03:01.056441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:20c37d9f4762256a64c2e20ad1f2f3744a756f1cf1ce798e7cdbd43d343552b1

Observation edbfce37-cd5d-40a1-aa97-70fcd4111e78 · inbound

A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration cites this paper.

A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:05:25.889426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T21:05:25.809179Z digest=sha256:6b902ebfcccf41f07730469308dda8f9981fa3f189f9233c05f74728935a0580

Observation a3538d8f-b61d-4dbc-a443-d79b5f05708b · inbound

GAIA: a benchmark for General AI Assistants cites this paper.

GAIA: a benchmark for General AI Assistants BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T15:46:03.502819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-12T15:46:03.247029Z digest=sha256:25ecc1613cccd197f539a84c95964dc9f7bd57080f1dcab2e0b1c9b944327bda

Observation 8ebc2ac2-b8ed-416f-ac32-27f70d54ce34 · inbound

WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks? cites this paper.

WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks? BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:48:05.210780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-15T02:48:04.972868Z digest=sha256:4e0af4759cd93337ea0e3736ec46ba325e349cb05bd2ba91bf0db01b4ee67e1e

Observation 50f349f0-7b3f-4d91-b0da-219db8d36a3d · inbound

A Survey on LLM-based Multi-Agent System: Recent Advances and New Frontiers in Application cites this paper.

A Survey on LLM-based Multi-Agent System: Recent Advances and New Frontiers in Application BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T05:29:24.509832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:29:24.509832Z digest=sha256:4db912cc35fa3c424007401e6b6a69df7aea8ab5d6b6451cd18de6f72eaef63d

Observation 4263410d-0eb7-435f-be7a-4fba1ea5f8c3 · inbound

NADER: Neural Architecture Design via Multi-Agent Collaboration cites this paper.

NADER: Neural Architecture Design via Multi-Agent Collaboration BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T00:52:22.895180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:52:22.895180Z digest=sha256:5e8d7be28de676b7a415ba5ff9ef7afa95fe179ce661521afa7806a3c33218ca

Observation 13708173-5c19-46fc-a319-8ed23e206cc8 · inbound

Flow: Modularized Agentic Workflow Automation cites this paper.

Flow: Modularized Agentic Workflow Automation BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:37:40.391435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:37:40.391435Z digest=sha256:2c86d068c865f48d9081bd5771df0e6febed3b200b03a4c7c39ab0aa3a773b1e

Observation d2e614de-7ab9-4b5c-aba0-95ad01828bf8 · inbound

Large Language Models for Multi-Robot Systems: A Survey cites this paper.

Large Language Models for Multi-Robot Systems: A Survey BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:32.203104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T04:32:05.138744Z digest=sha256:9260b7b526e6b1633e6e2c62db617b124bb0f42645bbc5fc0900966591f75e66

Observation 9fb02b18-1577-4654-9413-ef33ab7bbe61 · inbound

InstructRAG: Leveraging Retrieval-Augmented Generation on Instruction Graphs for LLM-Based Task Planning cites this paper.

InstructRAG: Leveraging Retrieval-Augmented Generation on Instruction Graphs for LLM-Based Task Planning BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T12:24:04.475783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:24:04.475783Z digest=sha256:edd0f7cac4232e5ed76b940344a48a99b36b19606c601ae527e3f00400eecabe

Observation 5af744e2-fe28-4738-b2fc-489e74b80782 · inbound

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization cites this paper.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.833363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:49.833363Z digest=sha256:1371a3e32a0095c68737ecf32a0c269177e7dac854235a2eceb546dc377e133b

Observation 6ce8b854-351e-4a55-a7be-178b2b272683 · inbound

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems cites this paper.

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:52:16.166118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T11:49:36.574471Z digest=sha256:667d1d740a9268b24380582cf95e646b742ab4049268dce450a719c2c13f6f85

Observation afb0e3e7-392c-4602-8f51-1b4de562688f · inbound

A Comprehensive Survey of Deep Research: Systems, Methodologies, and Applications cites this paper.

A Comprehensive Survey of Deep Research: Systems, Methodologies, and Applications BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 156

Resolution
unresolved
no resolver link, observed 2026-08-07T00:48:21.261348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:48:21.261348Z digest=sha256:e13b86f3f5d0f12514def7530534634a11c587ec6c09e394eea707ebd17b47a3

Observation f1500648-c6dc-4fb5-bec1-bc9d37bc47dd · inbound

MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models cites this paper.

MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:46:00.503290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:46:00.503290Z digest=sha256:3073763dc9881a3b9c78d3f89057dd2d9791dd6b6a92685a2bcc183297a6056d

Observation 6ca29658-beff-412c-a06f-e9c1d915008d · inbound

SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience cites this paper.

SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T23:55:49.795717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:55:49.795717Z digest=sha256:8f708204cc6b119669265fc9e605cfd577eb59e2f0a5cc367a90aa2e962d4737

Observation 18e22ced-d723-44c6-9338-3c010a917d43 · inbound

CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning cites this paper.

CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:50.685669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:50.685669Z digest=sha256:df1934b6fbedd944b41aef739258131f1e31afb674f6190b94aff23b6a9bcd56

Observation 55ebdc1b-1191-4e32-8427-e3c785c9369f · inbound

AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning cites this paper.

AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T16:08:39.320645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:08:39.320645Z digest=sha256:c4bc96f8fb9fd3056cc7015409fd495cc7cf78aac2a4452735ae2a67c01b825f

Observation 388fdae6-b527-463c-ae60-00610c7b0d83 · inbound

DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent cites this paper.

DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T19:46:40.814191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:46:40.814191Z digest=sha256:e26381831be72aa0bee6f1a9085d51314467b1d3e31108defc016c6724e5d295

Observation 009bcb12-5ac6-46d2-8cfc-f51b33d1f9b3 · inbound

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment cites this paper.

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:55:50.740383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-11T01:54:49.131461Z digest=sha256:6a134d51f7e7538a4ba7b03fa1669e11e963e24a031380689c76b3333d75cd4d

Observation fe3f45ff-5d7d-4726-85cb-06f99fbb9962 · inbound

LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability cites this paper.

LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 78

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T15:05:03.476844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-07-08T15:03:14.483228Z digest=sha256:f34abb4f24f0e24d7ff8443fe8ed8fc4c991beff84010e0dea1e5f961df22226

Observation 86314ec7-7baf-4d79-92e4-eaf93a573766 · inbound

Reproducing human biases in route choice using large language models: Toward scalable behavioral modeling cites this paper.

Reproducing human biases in route choice using large language models: Toward scalable behavioral modeling BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-14T04:15:33.637340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T04:15:33.637340Z digest=sha256:3d5d1749f3eebb5a511c858ae6f84bf99c1ac868503d2d5a68faa65a55f8e7ad

Observation 92516b54-3d9f-46f7-9a32-c128613e66bb · inbound

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details cites this paper.

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents

Reference 223

Resolution
unresolved
no resolver link, observed 2026-08-05T15:25:40.348435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:25:40.348435Z digest=sha256:efd6a3c3618f06d7c6bf96d002630936606302fbd8451e19932c67859ec5b200