Pith. sign in

Paper Citation Record · LEDGER

PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2411.00081.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.00081 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:42:30.589774Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T23:27:38.656546Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8a141125-b6a5-4702-8f69-3006b1a4fc4c · inbound

TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft cites this paper.

TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:53:34.919498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:53:34.919498Z digest=sha256:b756889c5117bb1e0d6df6436f94f34747c6c872586905b52ac27c3516d1b72d

Observation 5ffd778c-5ba8-43a5-82b1-5f5319a52ff6 · inbound

Effect of Adaptive Communication Support on LLM-powered Human-Robot Collaboration cites this paper.

Effect of Adaptive Communication Support on LLM-powered Human-Robot Collaboration PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T12:42:07.142284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:42:07.142284Z digest=sha256:070879bbfe494bba1d197809a24abacff764ebb6e1a29de815553e0123915948

Observation ba249ccf-2780-4f95-bdfb-a67ccd12da2c · inbound

ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding cites this paper.

ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T22:32:38.624156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:32:38.624156Z digest=sha256:6d1a8f5d806e2993f6c48945cba804b28c301bc6887007baa1a564a830f47ff3

Observation 05665e30-b543-419c-8c3a-a6438eb26726 · inbound

PLANET: A Collection of Benchmarks for Evaluating LLMs' Planning Capabilities cites this paper.

PLANET: A Collection of Benchmarks for Evaluating LLMs' Planning Capabilities PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:42:30.589774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:42:30.589774Z digest=sha256:5118c8ffb043e87def0de96aef36506471d1ca67c06a8c941d43af02c70bef0d

Observation 90fa97ae-9aba-45c6-a045-acbff35d598e · inbound

Collaborating Action by Action: A Multi-agent LLM Framework for Embodied Reasoning cites this paper.

Collaborating Action by Action: A Multi-agent LLM Framework for Embodied Reasoning PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T10:32:50.464554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:32:50.464554Z digest=sha256:a9336806737499fabde3a5254ab011e4236c533ca0afe7912ab7ed1c3078a413

Observation ae589b1e-e781-4a62-8341-cbddfde35355 · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:51.625209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:51.625209Z digest=sha256:e045e8a9d0215bf3677324553b6a6eae443e3cd28f0dfd617567625e67f83bbf

Observation 24701b72-183c-4e40-8b69-5a8f9fe7bd56 · inbound

TextAtari: 100K Frames Game Playing with Language Agents cites this paper.

TextAtari: 100K Frames Game Playing with Language Agents PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:57.934910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:57.934910Z digest=sha256:4781a6f6cbe31bdb6311095d0efc2560b247a83d74b53d2c906edcd897fc0bc0

Observation bc6a5c96-906c-4130-b6b0-394a37ef44cf · inbound

A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy cites this paper.

A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:51:55.370889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:51:55.370889Z digest=sha256:dfcfe9dbbc9db3689ae4d646c6fbc89a7e49157b25984478fa92a59a12b65720

Observation f9a3f08d-56b6-4d43-b786-a86940cb4b05 · inbound

Agent Identity Evals: Measuring Agentic Identity cites this paper.

Agent Identity Evals: Measuring Agentic Identity PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T14:57:06.679450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:57:06.679450Z digest=sha256:13b6308318b1887f2f2cc0a7068e70b3222d9d94a5683f369a4e660a325d03b2

Observation 1493e9ca-6552-4e83-9fd1-9439732b98df · inbound

ProToM: Promoting Prosocial Behaviour via Theory of Mind-Informed Feedback cites this paper.

ProToM: Promoting Prosocial Behaviour via Theory of Mind-Informed Feedback PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T05:43:44.535671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:43:44.535671Z digest=sha256:8ca5ef89975b5e0627bf4154a3f35226b5b9b3506a7f0147a7f01843e63c114e

Observation 3241174d-e951-4a3d-b54f-fcbe8d14bf2b · inbound

When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks cites this paper.

When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:01:08.966964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T08:59:35.944554Z digest=sha256:75d7d0e7f57479bd4be6e938b0064ca02a14cb0171905d17815da2048d4d84f3

Observation 254c8083-b32e-4293-b3a9-3f54ad28a658 · inbound

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data cites this paper.

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T08:15:14.592377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T08:15:14.592377Z digest=sha256:aa7ed87a0cda8b3d9bec2418ec106a6a7eb08e194dbd2db2d0b0633cb4d4d9a4

Observation d5a0eae3-e905-4563-add2-d0b6ca384061 · inbound

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs cites this paper.

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:00:40.476019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T06:00:02.043029Z digest=sha256:edf7b7b7f1f3e04e23ca4783e5685336991a186c69483e0efec7353bde9a5c19

Observation a6524d9d-dde1-4cbf-b589-9d0f485b7339 · inbound

Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks cites this paper.

Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T06:05:29.182557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:05:29.182557Z digest=sha256:6f2117e1f3ba8b074e94bd44fc7e4b43e7a10887134909e266f7ed446ceeb709

Observation a9a19770-539a-445b-a314-c2ad4cfe4685 · inbound

LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? cites this paper.

LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T22:26:34.782737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T22:26:34.782737Z digest=sha256:c57bbef30d4484d394a0a5d0e43d6d28e26ac1b414bdd9252936e085538a5894

Observation 2dca667a-62c8-4a17-8ad6-20fc1932113a · inbound

EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems cites this paper.

EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:06.677152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T15:21:20.231759Z digest=sha256:66abbfcb530a32735c98f263aac6f6659467480e42d0ead4452cdba8fcaca153

Observation c86b7f68-ac2e-40a8-9164-ed2d551725d5 · inbound

PersonalHomeBench: Evaluating Agents in Personalized Smart Homes cites this paper.

PersonalHomeBench: Evaluating Agents in Personalized Smart Homes PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:45:11.853525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T06:41:17.586880Z digest=sha256:db8c2bd965eced91b5d516d941383d7a763ee2a7ccc94374a39006590e863d48

Observation 41c036e8-2977-41c2-99cd-339ad2ed5e92 · inbound

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning cites this paper.

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:03:13.990666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T10:53:44.014259Z digest=sha256:4c9d9280f2eec90e77fbea21c5b22454fcf058acc5600b7e58837b02671cf9a9

Observation 97131752-cd1a-4e84-a37f-31322e583fd5 · inbound

Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models cites this paper.

Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:38:12.529366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T10:36:20.388169Z digest=sha256:016a47f3a6eec8549692f207818817c0bb4a3b3e3084bb93a9d616aeb02cdf60

Observation 3577e93f-7686-401b-9fca-57f7a887ab90 · inbound

PACT: Proactive Asking for Continual Task Assistance in Human-Robot Collaboration cites this paper.

PACT: Proactive Asking for Continual Task Assistance in Human-Robot Collaboration PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:44:41.298644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T13:34:46.547334Z digest=sha256:c691371038677bb9f40c97a79e5cfabfddd94802504ffc25997bf12f8a02e88e

Observation fe4cc417-286c-477a-a58a-84e8aebf2ca0 · inbound

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints cites this paper.

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T15:06:31.394162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:06:31.394162Z digest=sha256:38061d934ecbe7d9e85cd56e54ce42b4ea5a20ae38561a4b286877f222bf61a2

Observation b7900c57-6db3-472a-862d-15ca122da883 · inbound

LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior cites this paper.

LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:43:55.332960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T04:38:04.136601Z digest=sha256:9e14cc039b4d4bd0012a6627602c70b31b391ab0ce7eeb50b041b72cfd0d4873

Observation f760ef2d-dd2f-4aa0-87a2-7522e22190f2 · inbound

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models cites this paper.

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-08T02:44:27.711175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-08T02:44:00.608590Z digest=sha256:f6e90fa48f05899fe6648480da3daa9c4fbb0746a8d30041671eab335f8cdd1d

Observation abb6cf32-2618-412d-9753-8afbc5da791a · inbound

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models cites this paper.

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:13.133298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:13.133298Z digest=sha256:d4f0b0a5382372ecb3d2f940e96126fef157fbcd014cf90d2232d26b9d4dd898

Observation 9ad4e48c-1c38-47ba-80ed-fa30fb8871cc · inbound

CoMind: Understanding Collaborative Human Activity from Multiple Minds and Views cites this paper.

CoMind: Understanding Collaborative Human Activity from Multiple Minds and Views PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T23:27:38.672502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-10T23:20:12.870365Z digest=sha256:1c62d96af6fd4ded0939331c9b0b91e1969df4e754ef3405fd4f37c947b189f4

Observation d959edc4-cab0-491d-80f4-c34d7d274ea8 · inbound

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression cites this paper.

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T00:13:21.218440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:13:21.218440Z digest=sha256:e0b6f7bba3a0467915970e34d026fd270c7eac9044cc3fcc15588cc14f963bf1