Pith. sign in

Paper Citation Record · LEDGER

TravelPlanner: A Benchmark for Real-World Planning with Language Agents

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 53 inbound Pith citation observations for arXiv:2402.01622.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.01622 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 53 of 53 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:42:30.799753Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:45.572230Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f10056b9-1b38-4f5a-ac54-a089c3aca900 · inbound

Scaling Diffusion Language Models via Adaptation from Autoregressive Models cites this paper.

Scaling Diffusion Language Models via Adaptation from Autoregressive Models TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 194

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:59:36.686547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-20T19:59:36.461146Z digest=sha256:536520c6e12efe8e3eb2a72f2d51a9a38f5ef4e10b16594fe1ffaf3c33765437

Observation 96bb155f-b763-468d-9480-ca57fb65f362 · inbound

Large Language Models as User-Agents for Evaluating Task-Oriented-Dialogue Systems cites this paper.

Large Language Models as User-Agents for Evaluating Task-Oriented-Dialogue Systems TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T20:10:31.005888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:10:31.005888Z digest=sha256:e38f1aef7777eb385ff760e44662049b3a196109994f3935a4ec584c1fc5ca36

Observation be5e0c53-da8e-40b8-b390-bb4d6ab0354f · inbound

AI PERSONA: Towards Life-long Personalization of LLMs cites this paper.

AI PERSONA: Towards Life-long Personalization of LLMs TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T13:31:41.421416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:31:41.421416Z digest=sha256:0b52d55d033cdf2ae677f3a3cf6d3e93bdb87af49cd3baae2556b1bc494c8d4b

Observation 4d4f1d35-54c3-4654-8fb4-fb073d40067b · inbound

Evolving Deeper LLM Thinking cites this paper.

Evolving Deeper LLM Thinking TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T19:39:10.566537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:39:10.566537Z digest=sha256:8f59cb3599ba63cdf3486aea801229c6f7a68028865cbc53b4df692c19c742c3

Observation c24a3d8e-a0c3-49a3-aa91-b00ad7882174 · inbound

The AI Agent Index cites this paper.

The AI Agent Index TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-09T14:48:35.982122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:48:35.982122Z digest=sha256:a7ff617e31e29b37e49fa341bff31dc3ac09d20b52b13b0922d2c7d3700010aa

Observation a7db4069-f3a9-4ff8-9f2e-e3bd4ce01893 · inbound

AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting cites this paper.

AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:57:26.350724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-23T02:55:50.650423Z digest=sha256:c376186a88e1aff5feeff0c96f1f376ae929fbae3312946ccf7292bd453f1821

Observation 7a2634c9-b6a4-4f4c-a014-b3f983f6a537 · inbound

Large Language Model Agent: A Survey on Methodology, Applications and Challenges cites this paper.

Large Language Model Agent: A Survey on Methodology, Applications and Challenges TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:52:10.731335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T21:51:34.309870Z digest=sha256:83cbf7fa652787388a6976c658877457b07eca4a1df660f90347e994b22e6cef

Observation 613c7175-759a-412e-aa26-ba6f541983ef · inbound

PLANET: A Collection of Benchmarks for Evaluating LLMs' Planning Capabilities cites this paper.

PLANET: A Collection of Benchmarks for Evaluating LLMs' Planning Capabilities TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T11:42:30.799753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:42:30.799753Z digest=sha256:7b3e92745470d030a541026a3f48e1912feed9ee5cfcdcc4bf7c4a0c75703546

Observation c154d195-e624-463a-82d2-91ef604b4cbd · inbound

RLAP: A Reinforcement Learning Enhanced Adaptive Planning Framework for Multi-step NLP Task Solving cites this paper.

RLAP: A Reinforcement Learning Enhanced Adaptive Planning Framework for Multi-step NLP Task Solving TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T20:49:43.629393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:49:43.629393Z digest=sha256:dee4964fc70e3f4398ac5601851d9950e66dc709fea7e4e3c4ffa2e049f17bf0

Observation bd4733b0-9070-470a-b16c-076790d5aa46 · inbound

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation cites this paper.

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.825926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.825926Z digest=sha256:ea96c6c466e0c68c0e0ec1a6360a0427b340e48f0ea9c9bd1a438d786940d179

Observation c50db783-6e5d-466a-ad68-e9569395cfe1 · inbound

LLM-Powered AI Agent Systems and Their Applications in Industry cites this paper.

LLM-Powered AI Agent Systems and Their Applications in Industry TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 110

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:06:38.017908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T14:05:54.535411Z digest=sha256:252705e7bc65a88009c3aa369ab5476470e4806d7af5abad91826ea5a84214f0

Observation 40579da8-e258-4488-b850-eff0d6b85c0e · inbound

CoTGuard: Using Chain-of-Thought Triggering for Copyright Protection in Multi-Agent LLM Systems cites this paper.

CoTGuard: Using Chain-of-Thought Triggering for Copyright Protection in Multi-Agent LLM Systems TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:21.257734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:21.257734Z digest=sha256:8d78021e8a8427ea9109ea0845a7f49f7efe1727adaae1ddb48869c9a206cb0d

Observation c8972383-901a-467c-b332-3ca159221c5e · inbound

Alita: Generalist Agent Enabling Scalable Agentic Reasoning with Minimal Predefinition and Maximal Self-Evolution cites this paper.

Alita: Generalist Agent Enabling Scalable Agentic Reasoning with Minimal Predefinition and Maximal Self-Evolution TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:33.869420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:33.869420Z digest=sha256:503f5860abd0ac4834cf475de5eb40171cdd38721380fdbd5575033c0b578dcd

Observation 59645539-e3ca-4273-805c-8ce029e497e8 · inbound

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary cites this paper.

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:37:15.869784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T11:34:44.319579Z digest=sha256:4f30abc6d4b6b0e0aaed5a7232df31ed27e1e03f899ccd59b5df25c55c0d3a13

Observation 1b6eaf50-65fe-4888-aeed-bbb2b5a13b0d · inbound

Self-Challenging Language Model Agents cites this paper.

Self-Challenging Language Model Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:57.962659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:40:57.962659Z digest=sha256:694b4826f763ed98e46c42f74a52f81a7361cc584a00a65032b60467d1eeabc8

Observation f3260ffd-d064-4a19-9a25-ec8ac924f295 · inbound

Decompose, Plan in Parallel, and Merge: A Novel Paradigm for Large Language Models based Planning with Multiple Constraints cites this paper.

Decompose, Plan in Parallel, and Merge: A Novel Paradigm for Large Language Models based Planning with Multiple Constraints TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:23:20.212387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:23:20.212387Z digest=sha256:782fea37718992c561c7b19ee8f98fffc27c9af65246004a1a9c726cbfd6afe3

Observation 448a9493-dd86-438e-9ce2-ccd155fee5b5 · inbound

Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents cites this paper.

Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:05.443034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:05.443034Z digest=sha256:2e4e3eae3d085a269e419565c3e4a9824563ff085577ec57590c22a3d2801bb2

Observation 5fb0a280-7783-4f11-889f-dd962fea3173 · inbound

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning cites this paper.

MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:15:46.341641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:15:46.341641Z digest=sha256:dc7406452140144cd52134d5afe502ad2e51d50b476af0e5393432aedcf53062

Observation 8e0486c0-ab10-4aaf-a464-eb7c2fc8e9d9 · inbound

Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents cites this paper.

Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:44.505963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:36:44.505963Z digest=sha256:73649c89110a2befacba9ee85da2ceae8b847ef25a0986ebbbf6b6b7cc0916fa

Observation 322ba525-17be-48dd-8c64-a591647e86ab · inbound

Can Large Language Models Capture Human Risk Preferences? A Cross-Cultural Study cites this paper.

Can Large Language Models Capture Human Risk Preferences? A Cross-Cultural Study TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:54:43.768381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:54:43.768381Z digest=sha256:1a0694a88183ad548f6d166bbabe553969b17f24cc3ee6ef5b67dd22fd937456

Observation 4eeadb25-48fa-4d5b-ab1e-3591f80f59ef · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 210

Resolution
unresolved
no resolver link, observed 2026-08-06T17:54:17.900980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:54:17.900980Z digest=sha256:f0c27b5f4aa6662814361f18c33d8c26bc4e3f290a6e5b10365dc53b406563ac

Observation cd00d879-5d49-41dd-a89f-20c9d2c3469d · inbound

Hell or High Water: Evaluating Agentic Recovery from External Failures cites this paper.

Hell or High Water: Evaluating Agentic Recovery from External Failures TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T17:35:04.771353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:35:04.771353Z digest=sha256:e52e6d6c0bd20604a412bcc6b6d414829c784527ed91dad4786f81080a955f0f

Observation f1b793af-610f-4307-a96e-603d73518778 · inbound

Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems cites this paper.

Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T18:50:10.599269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:50:10.599269Z digest=sha256:61ed72bc06ea801b654599b754a7164ce7e9674385386cad3fb533c3d2d2a851

Observation 01350b3b-084e-410a-a850-b22e19f5f6fc · inbound

DoubleAgents: Human-Agent Alignment in a Socially Embedded Workflow cites this paper.

DoubleAgents: Human-Agent Alignment in a Socially Embedded Workflow TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:11:40.351710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T17:09:02.466082Z digest=sha256:69a05c62e8f26b9fe3868dabc9b7337dfa56058c2a817078172a06615d1566f7

Observation 7da19431-3761-4d6a-9971-6f077c944c11 · inbound

DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents cites this paper.

DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T15:49:16.077036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:49:16.077036Z digest=sha256:b5552a4c30992470b86cef006e066f3cef8ba84b6440c0966fc237522f08631b

Observation 7e643f6a-5480-4c1b-bb36-d0395c12efbc · inbound

When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks cites this paper.

When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:01:08.885097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T08:59:35.944554Z digest=sha256:0e20e4cf3a8900c25af776c7dfd1900d93137f3a96727c49d3535fb14a542ae9

Observation 034c5761-ddc3-4d71-86d4-b6c88d19c07f · inbound

COMPASS: Benchmarking Constrained Optimization in LLM Agents cites this paper.

COMPASS: Benchmarking Constrained Optimization in LLM Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T08:46:07.885854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T08:45:11.334594Z digest=sha256:88b87a193896ddae35ca6f1771e75178abdc3e5cf131b51836305b98a7d61f50

Observation a067f42e-c1a3-4e93-bf19-93c84b49f741 · inbound

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation cites this paper.

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:17:30.674636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T07:13:42.705619Z digest=sha256:3748591a6ab5a59ae335c3c52c00440c282f94084a2b498eda124e767f7e62c1

Observation 47840cab-ce8a-41f8-8f76-a98c2216def3 · inbound

LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? cites this paper.

LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T22:26:38.751311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T22:26:38.751311Z digest=sha256:378e46cc455119c6f7888a115b198f582c3453a249bbbc4dbbaa603244c27c54

Observation 04ff6d1f-ee53-41ec-8cb8-56a3ad70e7c9 · inbound

MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios cites this paper.

MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T20:41:10.606989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:41:10.606989Z digest=sha256:7bb1c8d89c91c32cc3b4e76994a617712c849916d335867c907cfa62226745e7

Observation 975aa7af-c2f7-42c6-ac05-186b8ec6c9a1 · inbound

HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents cites this paper.

HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T18:36:28.263051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T18:35:45.900606Z digest=sha256:a2b05b969f476d53f9e8e2078150c01ab4a03c2326bf2a105f4cef6e7ac75e5a

Observation 42632daa-e754-4765-9de1-a137ab952ba3 · inbound

Agentic AI for Trip Planning Optimization Application cites this paper.

Agentic AI for Trip Planning Optimization Application TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:31:19.457479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-09T19:45:44.908058Z digest=sha256:002d88202f90ed3212e19a7abde74f48d6461b3f384ee434f2b8c7ac51b88278

Observation c57135f6-cb62-40ee-abc1-d654cb9f10c1 · inbound

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning cites this paper.

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 122

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:38.987676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T18:19:55.849451Z digest=sha256:ae6a876074e3016dedf84658cc8441568edd7996bc8e14b5632a779bb0d9a8ed

Observation 0ead850d-da6c-4a36-8299-73b6b5695852 · inbound

TourMart: A Parametric Audit Instrument for Commission Steering in LLM Travel Agents cites this paper.

TourMart: A Parametric Audit Instrument for Commission Steering in LLM Travel Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:46:29.260041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T04:58:03.708461Z digest=sha256:3f43fb9afeb4c713a2ba40c1d3038a37011b5c8b369a8984d2972176be4bb81c

Observation 8e3fc6fd-faef-4075-a23f-86e25be2d3a5 · inbound

TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding cites this paper.

TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:28.000241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T04:49:03.891368Z digest=sha256:92f5e29abf709397bde79113ea1b712c36b5c72bba98d3aa40081342eb82d111

Observation 8fb6a719-fb2a-41e3-85a2-3b9f0a13dd1d · inbound

State-Centric Decision Process cites this paper.

State-Centric Decision Process TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:37:51.579903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-14T19:37:39.247986Z digest=sha256:16f25e77fbdc7cf3d5c937d862ff2cd81efd5fdcf9a7e646d324e5aee8bfbe7a

Observation 8eb446cb-9689-4966-8769-592bae2aea63 · inbound

Interactive Evaluation Requires a Design Science cites this paper.

Interactive Evaluation Requires a Design Science TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:13.970530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T10:55:08.135630Z digest=sha256:d8a9109757c91fe053c17d9d345e27d7bfc286e1955d62915f57b0b42f7f60de

Observation 8abf1099-102a-4670-8c9f-a4743613abf1 · inbound

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning cites this paper.

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:14:03.343839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T08:10:14.041279Z digest=sha256:9b076abb15144068cfc0306a30561ffe5858564f731e9565ac6364673f5a0bd5

Observation 46bbace4-1160-4249-9603-213313f6a2a4 · inbound

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning cites this paper.

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:55:48.262269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T18:38:09.609972Z digest=sha256:689dceaf489d1e3ce45b2df5dc01fd8a4271f5e5dfb1cd7b9693648b46763c26

Observation 0847ff6b-6d0c-4c85-a989-05994e3490fb · inbound

TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents cites this paper.

TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:22:24.595985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T17:20:29.096427Z digest=sha256:ec50b181a7dc404aa4b56bab8e0a529439eb5773f08d1da4dc3207c684eb2cc5

Observation f4483946-daa2-4554-93bd-0b7e977320f8 · inbound

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline cites this paper.

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:36:45.437704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T06:49:08.689588Z digest=sha256:273ba7a32dc6ff486b7a7e0867a736dbd4c5451fd3dd646641867f435edbcb7a

Observation 65c6171e-292b-4d53-aa05-668606b84fe3 · inbound

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints cites this paper.

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T15:06:31.394162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:06:31.394162Z digest=sha256:7d6665009ce2fc23b4213d0bf00e58df6a57d90b4cd3e85efe1e876857eb6162

Observation a2aaa4e3-1929-4c9e-8278-4de0f0dcdd7f · inbound

OPENPATH: A Supervisor--Specialist Agent System for Personalized, Accessible, and Multi-stop Urban Trip Planning cites this paper.

OPENPATH: A Supervisor--Specialist Agent System for Personalized, Accessible, and Multi-stop Urban Trip Planning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:07:22.398862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T20:53:38.995053Z digest=sha256:55ab2eac5784ff1b9186b168dc97430b79244c4e3fa06edb34a5b4cdba8cb6a4

Observation 51759abd-6529-4bc8-96f3-6c5139d4629c · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 183

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.889153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:3c3143f502a19989466558c52c28bfa81fbcfbec5508e6d6b0a4cc850f85a90a

Observation 75da41b7-fb71-44be-8ac9-8baa282f61c6 · inbound

REVES: REvision and VErification--Augmented Training for Test-Time Scaling cites this paper.

REVES: REvision and VErification--Augmented Training for Test-Time Scaling TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:29:15.117250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T21:14:15.337979Z digest=sha256:2a83a79e65ec47423813b3679c96706f67dee5503ce5e39deab812c52335aae9

Observation 076eba6c-9c5a-4831-a534-a9b74b6a0aa0 · inbound

StaminaBench: Stress-Testing Coding Agents over 100 Interaction Turns cites this paper.

StaminaBench: Stress-Testing Coding Agents over 100 Interaction Turns TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:19:23.936321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T19:47:59.090249Z digest=sha256:5b15296701b37a0b9975edbbe308216bbc0c531332bba70d5dd46d428c182d1c

Observation 4052e4cf-6822-4ce0-9e00-d980d376bae2 · inbound

Trip+: Benchmarking Agents in Personalized Interactive Travel Planning cites this paper.

Trip+: Benchmarking Agents in Personalized Interactive Travel Planning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:19:37.942696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T14:38:38.844555Z digest=sha256:3a2f5492e70b841c64d318ba702d2338e41e08124310d77a85bf0e4bee3d0523

Observation 5e66c830-d13f-4d8e-8f42-2ae3a9868239 · inbound

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems? cites this paper.

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems? TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:45.574347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T09:15:50.722199Z digest=sha256:e687e22c0e6086f2e4b92fd1d27da6d64a682373ab82c7c93eae9a445b002c22

Observation 3493e061-3254-4da5-98f2-2067343e4c13 · inbound

When Do Multi-Agent Systems Help? An Information Bottleneck Perspective cites this paper.

When Do Multi-Agent Systems Help? An Information Bottleneck Perspective TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T21:18:07.731086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:18:07.731086Z digest=sha256:1d8213cabaf439399d11ff0611821e1ea320fe84832e5caced04e0760ed6852d

Observation e87b2bab-6920-43fb-ae3f-e28ea6f97489 · inbound

AlterAtlas: Shifting Travel Planning from AI Generation to Validation via Persona-Driven Simulations cites this paper.

AlterAtlas: Shifting Travel Planning from AI Generation to Validation via Persona-Driven Simulations TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T20:37:59.709268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:37:59.709268Z digest=sha256:9cc6c0f711c6828e00e45baeedc741b07c07f4c5140a2847596bbb77f0e998db

Observation 745aafea-9d2b-4b92-aa44-197201019630 · inbound

ProEvent: An Event-centric Benchmark for Proactive Agents cites this paper.

ProEvent: An Event-centric Benchmark for Proactive Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-01T17:18:08.072328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:18:08.072328Z digest=sha256:377289e1b9a3f80e30c23b20637aa8555fd2eb0b26f8c24f12b5e5206172b943

Observation 0672470c-508c-4b1d-80eb-51b4acec5069 · inbound

Think2Go: Generative Next POI Recommendation with LLM Reasoning cites this paper.

Think2Go: Generative Next POI Recommendation with LLM Reasoning TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T15:48:02.235323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:48:02.235323Z digest=sha256:b3bfbcbb79893045af18b2c14c9dda48f24aca1ded6710d821c73e8baa748675

Observation 8abd30ba-20c7-43a7-b0c4-80799e381cdd · inbound

LatticeMind: A Conflict-Aware Memory Primitive for Multi-Agent Systems cites this paper.

LatticeMind: A Conflict-Aware Memory Primitive for Multi-Agent Systems TravelPlanner: A Benchmark for Real-World Planning with Language Agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T00:19:26.672694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:19:26.672694Z digest=sha256:1c4e9ad9ddb4d98ffda9cf34e021df8bb5d55acfd908af8209d09c410a7e8f32