Pith. sign in

Paper Citation Record · LEDGER

WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2405.00823.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.00823 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:35:50.618120Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 17eb5e90-6760-475f-b74f-04a88d54dcc2 · inbound

Kaleidoscopic Teaming in Multi Agent Simulations cites this paper.

Kaleidoscopic Teaming in Multi Agent Simulations WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:50.618120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:50.618120Z digest=sha256:ea928276842a4a413cbdf77d6899b019560585beb5185022c44e9cc13c3ce592

Observation 8c1c9a02-0242-4508-b36f-1b5b7a543002 · inbound

Benchmarking Deep Search over Heterogeneous Enterprise Data cites this paper.

Benchmarking Deep Search over Heterogeneous Enterprise Data WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:53:43.371878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:53:43.371878Z digest=sha256:18d24697287bdea09b63a416631e6e9f8889a9651bc69f13ac3d606eee6bdd8b

Observation 165a74fe-c443-4b78-ba82-a10d714665cd · inbound

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems cites this paper.

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:21:42.700042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T23:21:42.029285Z digest=sha256:4813db0dc0e8b666185d8a9445261a3f9928ffb751a5b2295adfd91b9c48ae50

Observation 456a1879-7af9-4cb3-84f3-e8c05b41cda1 · inbound

Memory in the Age of AI Agents cites this paper.

Memory in the Age of AI Agents WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 297

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:18:20.504853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T18:18:19.911342Z digest=sha256:519f72ba7963ea683e576ac3b06190a50673b24b3d3d2b31fa0824f47518ee8e

Observation c0f85bf0-b868-4009-b8c8-245a28c4dc56 · inbound

Gecko: A Simulation Environment with Stateful Feedback for Refining Agent Tool Calls cites this paper.

Gecko: A Simulation Environment with Stateful Feedback for Refining Agent Tool Calls WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T21:46:33.369431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:46:33.369431Z digest=sha256:c114b0a042a6bcf88f23d206abb15203f8f9aeb3df880494a7f94844869a328d

Observation 787d2f58-fd9a-420d-9da5-36c7f9716047 · inbound

TacoMAS: Test-Time Co-Evolution of Topology and Capability in LLM-based Multi-Agent Systems cites this paper.

TacoMAS: Test-Time Co-Evolution of Topology and Capability in LLM-based Multi-Agent Systems WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:24.582334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T04:53:54.754878Z digest=sha256:e09803350724bbd99ec2358f10c63e5ba14c364eea03d91507e3c0833455f13d

Observation c955e6bb-26ca-4fad-b4bc-c6f5b16cf412 · inbound

Agents that Matter: Optimizing Multi-Agent LLMs via Removal-Based Attribution cites this paper.

Agents that Matter: Optimizing Multi-Agent LLMs via Removal-Based Attribution WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T14:43:30.891950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T14:36:25.320506Z digest=sha256:8a59c7811ef8cc5ea7f0c6adfb681bee046174c485d13ba7e398ff2a0e9e6d90

Observation 890132bf-5411-4278-bece-7f092c012ba6 · inbound

MemToolAgent: Leveraging Memory for Tool Using Agents Based on Environment and User Feedback cites this paper.

MemToolAgent: Leveraging Memory for Tool Using Agents Based on Environment and User Feedback WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:37:22.830276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:17:36.472178Z digest=sha256:d3b81fb0761ca668f96e7507436993c3e654e90b6d7753232e6295d021070c42

Observation 0232d706-fa8a-42c3-a4a9-c1efc5528051 · inbound

Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation cites this paper.

Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-26T08:49:15.272343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T08:39:41.575494Z digest=sha256:46a8a2f1788260430afd4324bd6384cb9a7b62f4a12fa25789e334f45c84a575

Observation 8e62aba3-56c5-457a-8666-788914c0b110 · inbound

When Do Multi-Agent Systems Help? An Information Bottleneck Perspective cites this paper.

When Do Multi-Agent Systems Help? An Information Bottleneck Perspective WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T21:18:07.245600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:18:07.245600Z digest=sha256:e6d165a97036e2e8d6253f3ae13cb53631f0588d113fa304fd5c39e55676e413

Observation a940c768-4480-4847-903c-17a989b7c168 · inbound

Scores Are Not Decisions: Cost-Aware Stopping for Tool Acquisition in LLM Agents cites this paper.

Scores Are Not Decisions: Cost-Aware Stopping for Tool Acquisition in LLM Agents WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-30T11:41:09.738070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T11:41:09.738070Z digest=sha256:39fee3c301d005e9cf29d8ff12c9f124c017df4b30fc1b5e10e5e8e1bf10ad65

Observation 092207ef-8fef-46f4-ad6b-2ef6e2675dbe · inbound

GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks cites this paper.

GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T13:09:08.945549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:09:08.945549Z digest=sha256:c00bb4828cb8780121c0fe5a2a1048494c8ef3c10a58497b97b38e4fa7ab7da6