Pith. sign in

Paper Citation Record · LEDGER

TravelPlanner: A Benchmark for Real-World Planning with Language Agents

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2402.01622.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.01622 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:27.825926Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:45.572230Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f10056b9-1b38-4f5a-ac54-a089c3aca900 · inbound

Scaling Diffusion Language Models via Adaptation from Autoregressive Models cites this paper.

Scaling Diffusion Language Models via Adaptation from Autoregressive Models TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 194

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:59:36.686547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T19:59:36.461146Z digest=sha256:29c3fe709e9a4975840fa30ec9a219d0d8406bd3799dfa4ee7a4e3f503f3a48f

Observation a7db4069-f3a9-4ff8-9f2e-e3bd4ce01893 · inbound

AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting cites this paper.

AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:57:26.350724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-23T02:55:50.650423Z digest=sha256:3f16435b65e9f9474a31f4c72a529c25e41a4e7aea9d5cd50ff94e9fbe53fb05

Observation 7a2634c9-b6a4-4f4c-a014-b3f983f6a537 · inbound

Large Language Model Agent: A Survey on Methodology, Applications and Challenges cites this paper.

Large Language Model Agent: A Survey on Methodology, Applications and Challenges TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:52:10.731335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T21:51:34.309870Z digest=sha256:a06636e2177e5ce91f58e01ed9835ac2433eed657707092fe00e0389cd18cd91

Observation bd4733b0-9070-470a-b16c-076790d5aa46 · inbound

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation cites this paper.

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.825926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.825926Z digest=sha256:3e6cb978ffe1a1b65cd1336264cd950dc15b649478f7f7253322c87def585b8a

Observation c50db783-6e5d-466a-ad68-e9569395cfe1 · inbound

LLM-Powered AI Agent Systems and Their Applications in Industry cites this paper.

LLM-Powered AI Agent Systems and Their Applications in Industry TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 110

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:06:38.017908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T14:05:54.535411Z digest=sha256:0f082180f84e3face069f9ab77f21401418b91dd5c2064989cca20e86e2b1f37

Observation 40579da8-e258-4488-b850-eff0d6b85c0e · inbound

CoTGuard: Using Chain-of-Thought Triggering for Copyright Protection in Multi-Agent LLM Systems cites this paper.

CoTGuard: Using Chain-of-Thought Triggering for Copyright Protection in Multi-Agent LLM Systems TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:21.257734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:21.257734Z digest=sha256:a1462d4d8892aece75ded4d0a73dcf3921582a6d6f91130956a5447a42f7e9cb

Observation c8972383-901a-467c-b332-3ca159221c5e · inbound

Alita: Generalist Agent Enabling Scalable Agentic Reasoning with Minimal Predefinition and Maximal Self-Evolution cites this paper.

Alita: Generalist Agent Enabling Scalable Agentic Reasoning with Minimal Predefinition and Maximal Self-Evolution TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:33.869420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:33.869420Z digest=sha256:b798ddf343680e2dd0a00355607c3746b3f200ed83da615ff8bbb32a2605d64b

Observation 59645539-e3ca-4273-805c-8ce029e497e8 · inbound

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary cites this paper.

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:37:15.869784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T11:34:44.319579Z digest=sha256:c6723795498a6e06b82212d7dc650e211c817355981a5694c1ebdbea256e65c0

Observation 1b6eaf50-65fe-4888-aeed-bbb2b5a13b0d · inbound

Self-Challenging Language Model Agents cites this paper.

Self-Challenging Language Model Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:57.962659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:40:57.962659Z digest=sha256:45b096828390cc2e7c1feca51d3b63e353662e086c3818d283165ae62ce16236

Observation f3260ffd-d064-4a19-9a25-ec8ac924f295 · inbound

Decompose, Plan in Parallel, and Merge: A Novel Paradigm for Large Language Models based Planning with Multiple Constraints cites this paper.

Decompose, Plan in Parallel, and Merge: A Novel Paradigm for Large Language Models based Planning with Multiple Constraints TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:23:20.212387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:23:20.212387Z digest=sha256:bcf09617f8639ee3504f38d7b60295883b6d2313f4f1786ca8d99bc8c0fc07a1

Observation 448a9493-dd86-438e-9ce2-ccd155fee5b5 · inbound

Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents cites this paper.

Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:05.443034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:05.443034Z digest=sha256:8a5c13d746e3331ab42b8ea80fa013c0c8283c4c36a0ead0cc5653280055cf38

Observation 5fb0a280-7783-4f11-889f-dd962fea3173 · inbound

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning cites this paper.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.341641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.341641Z digest=sha256:32c5656b78ced9eb208ad84e9512948662e85ed8199e0b1888c146264adedad2

Observation 8e0486c0-ab10-4aaf-a464-eb7c2fc8e9d9 · inbound

Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents cites this paper.

Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:44.505963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:36:44.505963Z digest=sha256:b36d4d771ece3879a3fc7793eabcdd87cda86446c9c61e5dc7321c89a296239a

Observation 322ba525-17be-48dd-8c64-a591647e86ab · inbound

Can Large Language Models Capture Human Risk Preferences? A Cross-Cultural Study cites this paper.

Can Large Language Models Capture Human Risk Preferences? A Cross-Cultural Study TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:54:43.768381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:54:43.768381Z digest=sha256:9e1c81e2522dcbf41e9b07f8a1ae4f4590e8976d4d93174af5cd75a99c7896fc

Observation 4eeadb25-48fa-4d5b-ab1e-3591f80f59ef · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 210

Resolution
unresolved
no resolver link, observed 2026-08-06T17:54:17.900980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:54:17.900980Z digest=sha256:3c37c5b6bc511088bf955d87c2033b597d93939a5e2721cb55895a6909ca8990

Observation f1b793af-610f-4307-a96e-603d73518778 · inbound

Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems cites this paper.

Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T18:50:10.599269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:50:10.599269Z digest=sha256:51ee7b9f10699a59e5def3f85c2d470c5e1a61bc1a90e214d0c561af58ac5e2c

Observation 01350b3b-084e-410a-a850-b22e19f5f6fc · inbound

DoubleAgents: Human-Agent Alignment in a Socially Embedded Workflow cites this paper.

DoubleAgents: Human-Agent Alignment in a Socially Embedded Workflow TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:11:40.351710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T17:09:02.466082Z digest=sha256:e754b5ba33b7daa9e3fbe75c4bb3f1dc14746299bc98a8488d1b29e1825e1e10

Observation 7e643f6a-5480-4c1b-bb36-d0395c12efbc · inbound

When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks cites this paper.

When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:01:08.885097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T08:59:35.944554Z digest=sha256:76a4149b65792fb45def34cab53e960a2e1efe5eb2296195042f6fa7863c56a1

Observation 034c5761-ddc3-4d71-86d4-b6c88d19c07f · inbound

COMPASS: Benchmarking Constrained Optimization in LLM Agents cites this paper.

COMPASS: Benchmarking Constrained Optimization in LLM Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T08:46:07.885854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T08:45:11.334594Z digest=sha256:2ca13fc4bd54b096502f62bb10c72f528f1f969c696fb8167a39ef0688266929

Observation a067f42e-c1a3-4e93-bf19-93c84b49f741 · inbound

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation cites this paper.

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:17:30.674636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T07:13:42.705619Z digest=sha256:4d0687980dcf8e89c5a14c9ea65a1c100e2f49e7538a1980823dd91982c4936c

Observation 47840cab-ce8a-41f8-8f76-a98c2216def3 · inbound

LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? cites this paper.

LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T22:26:38.751311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T22:26:38.751311Z digest=sha256:b19e4e537dc8b87a7a61e402e1c35c837b56d45a39f70dc6108b6a02b224ab4c

Observation 04ff6d1f-ee53-41ec-8cb8-56a3ad70e7c9 · inbound

MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios cites this paper.

MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T20:41:10.606989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:41:10.606989Z digest=sha256:0a470c0cd0418a323f084c8a4e0cdf690a8d2e22d1163ca556c38fbea7258286

Observation 975aa7af-c2f7-42c6-ac05-186b8ec6c9a1 · inbound

HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents cites this paper.

HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T18:36:28.263051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T18:35:45.900606Z digest=sha256:f34272906c0bf427509625a6ddadfdf55e5d6633da75892bfd068542ea5ff0b9

Observation 42632daa-e754-4765-9de1-a137ab952ba3 · inbound

Agentic AI for Trip Planning Optimization Application cites this paper.

Agentic AI for Trip Planning Optimization Application TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:31:19.457479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T19:45:44.908058Z digest=sha256:821d3a4efe905c9a34fbbe6d48ac9cd3145b221589ddfa6993df89b35e39bbec

Observation c57135f6-cb62-40ee-abc1-d654cb9f10c1 · inbound

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning cites this paper.

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 122

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:38.987676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T18:19:55.849451Z digest=sha256:b3c69f7fa425409fc5ea8ecf3a1cab5b6ad2d379e26bca5d0a708548ef04a7fc

Observation 0ead850d-da6c-4a36-8299-73b6b5695852 · inbound

TourMart: A Parametric Audit Instrument for Commission Steering in LLM Travel Agents cites this paper.

TourMart: A Parametric Audit Instrument for Commission Steering in LLM Travel Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:46:29.260041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T04:58:03.708461Z digest=sha256:245696b7d379289bad5b9cafd12a7fb7a1764cabe896a8910256da59c62399ff

Observation 8e3fc6fd-faef-4075-a23f-86e25be2d3a5 · inbound

TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding cites this paper.

TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:28.000241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T04:49:03.891368Z digest=sha256:63c0f63fe7b675481fbd6546a6bc0cd66fdc3e2f2ab9e3be8ea017c78fb61010

Observation 8fb6a719-fb2a-41e3-85a2-3b9f0a13dd1d · inbound

State-Centric Decision Process cites this paper.

State-Centric Decision Process TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:37:51.579903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T19:37:39.247986Z digest=sha256:0664c385c2b4f814820ff7ecca77fb795ec53cb03fedf42ec81cfa1a9e4ba350

Observation 8eb446cb-9689-4966-8769-592bae2aea63 · inbound

Interactive Evaluation Requires a Design Science cites this paper.

Interactive Evaluation Requires a Design Science TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:13.970530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:55:08.135630Z digest=sha256:d4d4f056c3a2bb6ecc72a75147c07e5dbf0c561fd043661d8c0c1b5112459f82

Observation 8abf1099-102a-4670-8c9f-a4743613abf1 · inbound

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning cites this paper.

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:14:03.343839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T08:10:14.041279Z digest=sha256:b1488652d405ae34c55a516542a73e26de5d0ea0700e206bc964ab842d611efc

Observation 46bbace4-1160-4249-9603-213313f6a2a4 · inbound

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning cites this paper.

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:55:48.262269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T18:38:09.609972Z digest=sha256:2bfd0ef94ba290d068bd4cdacbe0ec116900e42148d75591e7e9880734ee8133

Observation 0847ff6b-6d0c-4c85-a989-05994e3490fb · inbound

TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents cites this paper.

TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:22:24.595985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T17:20:29.096427Z digest=sha256:afe5f89fe40df6596d62350dcd518826cd890980c541c0ecdaba9ddd3252d73d

Observation f4483946-daa2-4554-93bd-0b7e977320f8 · inbound

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline cites this paper.

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:36:45.437704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T06:49:08.689588Z digest=sha256:f362c7577d374d6a0d5a78a13d299cc3ea3177d3219eac2839cd69563eadbb5e

Observation 65c6171e-292b-4d53-aa05-668606b84fe3 · inbound

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints cites this paper.

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T15:06:31.394162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:06:31.394162Z digest=sha256:db155657bbdbb97338e61ef09e04d503cf4e4d52cb95c3795c4b9e00a646c56b

Observation a2aaa4e3-1929-4c9e-8278-4de0f0dcdd7f · inbound

OPENPATH: A Supervisor--Specialist Agent System for Personalized, Accessible, and Multi-stop Urban Trip Planning cites this paper.

OPENPATH: A Supervisor--Specialist Agent System for Personalized, Accessible, and Multi-stop Urban Trip Planning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:07:22.398862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T20:53:38.995053Z digest=sha256:6fd5a3336d620f6edea26c73ce7594ff31815b5ca8a0fd412f53562940fd708e

Observation 51759abd-6529-4bc8-96f3-6c5139d4629c · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 183

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.889153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:8d67a627961384406b9898130d03b87b91c093753cc1a80a4e46cb1cbc50435f

Observation 75da41b7-fb71-44be-8ac9-8baa282f61c6 · inbound

REVES: REvision and VErification--Augmented Training for Test-Time Scaling cites this paper.

REVES: REvision and VErification--Augmented Training for Test-Time Scaling TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:29:15.117250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T21:14:15.337979Z digest=sha256:97b7659cf4ed298f22d8f0c42ca6f89b0ad98efe3bcc3e8a587e83922bad362f

Observation 076eba6c-9c5a-4831-a534-a9b74b6a0aa0 · inbound

StaminaBench: Stress-Testing Coding Agents over 100 Interaction Turns cites this paper.

StaminaBench: Stress-Testing Coding Agents over 100 Interaction Turns TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:19:23.936321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T19:47:59.090249Z digest=sha256:cec8813e97960b3ddb17fad1ea75d28d8d0803f2118619770a1ecb98fa6718fc

Observation 4052e4cf-6822-4ce0-9e00-d980d376bae2 · inbound

Trip+: Benchmarking Agents in Personalized Interactive Travel Planning cites this paper.

Trip+: Benchmarking Agents in Personalized Interactive Travel Planning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:19:37.942696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T14:38:38.844555Z digest=sha256:61ffc356d25a96edcd50b813d7cbd663f58bd4188b67afc775936462d0aeedbb

Observation 5e66c830-d13f-4d8e-8f42-2ae3a9868239 · inbound

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems? cites this paper.

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems? TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:45.574347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T09:15:50.722199Z digest=sha256:13874e38437bc184e2acd7364534b59476891696efb731f664bcf1631cd7cfe5

Observation 3493e061-3254-4da5-98f2-2067343e4c13 · inbound

When Do Multi-Agent Systems Help? An Information Bottleneck Perspective cites this paper.

When Do Multi-Agent Systems Help? An Information Bottleneck Perspective TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T21:18:07.731086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:18:07.731086Z digest=sha256:f779562e15db35f20b6984507b929adb503ab8b2d77984ca2cfac849c84bd076

Observation e87b2bab-6920-43fb-ae3f-e28ea6f97489 · inbound

AlterAtlas: Shifting Travel Planning from AI Generation to Validation via Persona-Driven Simulations cites this paper.

AlterAtlas: Shifting Travel Planning from AI Generation to Validation via Persona-Driven Simulations TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T20:37:59.709268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:37:59.709268Z digest=sha256:d1f3a40fdbb3acce91ebd088dc1db513a7c5dfa810184213c0f4a44c18c8a577

Observation 745aafea-9d2b-4b92-aa44-197201019630 · inbound

ProEvent: An Event-centric Benchmark for Proactive Agents cites this paper.

ProEvent: An Event-centric Benchmark for Proactive Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-01T17:18:08.072328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:18:08.072328Z digest=sha256:a8a2e6fd11eae27e035335f941be990f293d48ef782685eda4f6e73c4520551b

Observation 0672470c-508c-4b1d-80eb-51b4acec5069 · inbound

Think2Go: Generative Next POI Recommendation with LLM Reasoning cites this paper.

Think2Go: Generative Next POI Recommendation with LLM Reasoning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T15:48:02.235323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:48:02.235323Z digest=sha256:00cd42cabed482a315b5d3ac1f356451735cc7f269874ab6407f5fa9c3650ba0