Pith. sign in

Paper Citation Record · LEDGER

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

As of 5 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2606.11816.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.11816 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T09:47:04.122464Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact16
  • verified fuzzy0
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 76230b58-d6e9-4103-9c32-2283fd57a154 · outbound

This paper cites 2022 , booktitle=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2022 , booktitle=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:6ec142bfe26f3440c7db5c2c90a68322662ed565c9750cd689ccab01ed063fdf

Observation d78c7011-b4d5-4f58-a32b-15396f2b7535 · outbound

This paper cites 2025 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2025 , eprint=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:747278cb5470cd35e65c29f84d9852c3765bd81d5aa83894cabda2164ba8da73

Observation 41fc2712-91e5-4a64-8b56-1789e12b1f16 · outbound

This paper cites 2023 , booktitle=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2023 , booktitle=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:49fdb111731097ffed64cc62f851dee470a6c62b28a9c1216d9f85a036efd67c

Observation 0f47a672-d2dd-4056-bf7c-8c653c8f6227 · outbound

This paper cites 2025 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2025 , eprint=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:072129d39eb5b576df997147dfd40669518bc4d1319057dbaed27301289d5691

Observation cf47892a-37ca-4f3c-8956-5ca778121a45 · outbound

This paper cites 2025 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2025 , eprint=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:7c26d3aa5eb7e64f227cdd33631bef0118735b04f248993b0ac741c624d4d909

Observation 4073cd05-cb24-4a3c-af0c-cb34462874cb · outbound

This paper cites 2025 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2025 , eprint=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:7f75f35fbf4c760bf14320906bf65f71e5317b607481334a5ff38b6ba124d3f9

Observation 13c9efb7-3b9d-4ce7-b377-54b15bc56fec · outbound

This paper cites 2024 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2024 , eprint=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:084b9f1feff02cd99109eb90f66be0ccec5adf89710dc49d98bfb3731552d5c5

Observation 88bf4e0d-5d4d-4f5b-855f-220dc1cb45ad · outbound

This paper cites 2025 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2025 , eprint=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:05b1b52c4a3afa0a1ea9ab6db263faa94006a6d88c7072c5e5ce21b85045b5e9

Observation 8053b767-8a20-4daa-a528-74d02f51e9cf · outbound

This paper cites 2024 , booktitle=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2024 , booktitle=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:2be46d7202d17537afcd46fdda0a21ec496c5a6fcae0850545ad34d6b7b739be

Observation 893e4bb9-00ca-401b-bac4-3dcd2351d030 · outbound

This paper cites 2024 , booktitle=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2024 , booktitle=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:ba5c415b0467b61c735b125fd198188b7f9b17dad57dd4d6faf4aaa377064d5e

Observation 35488357-89a7-4945-aecd-ee62aeb8ffe0 · outbound

This paper cites 2024 , booktitle=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2024 , booktitle=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:3f9dfd973ee54beb8e6dd461092009e01d4c99c837740edff1cdd68e159167ba

Observation 1cf99f21-8b31-4f5e-ae84-8a4f7a51c0e7 · outbound

This paper cites 2025 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2025 , eprint=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:52ece3b2054607fd6d3059fdb56daeb3a2f525edf71ed04e94e9d0556388b3d4

Observation 6ffe18b6-c5b3-4afa-9fc7-a96e62c37521 · outbound

This paper cites 2024 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2024 , eprint=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:f2e9693f2b0b438ffb1ad073a5babe6d78bddb687d86dab09d60fbbdad170259

Observation d8bee4f1-ba96-496c-a1c5-16d2d44ab4fb · outbound

This paper cites Journal of Behavioral Decision Making , volume=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Journal of Behavioral Decision Making , volume=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:e01073d5c1f255e3e314e890523e67cb2283d0eddf36731c162c7ac207e005e1

Observation 52e7e8e3-bcee-464d-bdde-48eeed1dfda1 · outbound

This paper cites Ramnani and Mayuresh Anand and Shubhashis Sengupta and Andrew E.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Ramnani and Mayuresh Anand and Shubhashis Sengupta and Andrew E

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:83012a16c0373320fe2ef7377916198a9f052cde6787ad753a51474ee8e06d97

Observation e28d7a81-72f1-4a25-944e-503f86bc8e0d · outbound

This paper cites 2023 , booktitle=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2023 , booktitle=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:535c6a4e5fc78010402792c19cad0d33cb7dc1818635a987db31e446dc01d110

Observation 898d2ac3-3c4f-43cd-8db7-aa7a6dac8c35 · outbound

This paper cites 2022 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2022 , eprint=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:cdf2b7bde240ecea53e20cef722a9917bb956b9c287a999e9427bb5be6d89e60

Observation 29d3ecbf-a376-4432-a3c2-dfe82d258560 · outbound

This paper cites an unresolved cited work.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:8b768cea571aa152c71c4438c2702e806bd2a194c9c8ae08da6d2965538b3b00

Observation 9622fe35-1be2-414b-85db-5b82a7f12615 · outbound

This paper cites 2025 , eprint=.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning 2025 , eprint=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:f14da7813ffdf0e5598b8715965363dd8b3c369999f31ee0603e3b6db7f2c638

Observation 286b0ba0-0aae-44c6-a48d-f7a365d1c9ab · outbound

This paper cites , booktitle =.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning , booktitle =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:ae2df191eaf56290d3a88f81917ad2f9b469e1400b6d711b7c2639c9b9a9af9e

Observation d10e56ba-7de7-496e-b7ad-6f43cb5879cf · outbound

This paper cites ForecastTKGQuestions: A Benchmark for Temporal Question Answering and Forecasting over Temporal Knowledge Graphs.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning ForecastTKGQuestions: A Benchmark for Temporal Question Answering and Forecasting over Temporal Knowledge Graphs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:83db3ae3d84b74dff49e0d45a768625d520ed7ffa64993ef2963cd4d3e8af6ab

Observation d98000fa-a14c-48b2-bc84-bf61a31f1c8e · outbound

This paper cites BERT : Pre-training of Deep Bidirectional Transformers for Language Understanding.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning BERT : Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 25

Resolution
verified exact
doi, observed 2026-06-27T09:50:48.137102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:72c56a61380dfd395759bf3beec9440a1611422c9266ba9fec054ca450ef21bc

Observation ce43e413-c15f-495c-8d1f-ef1483a0605c · outbound

This paper cites an unresolved cited work.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:d2135fd35f42f763b195469e797517c0374c75dbd80ff4743cf9d28735e77344

Observation 83620d71-1191-41d7-81c9-2878bd9c5b34 · outbound

This paper cites Are LLMs Prescient? A Continuous Evaluation using Daily News as the Oracle.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Are LLMs Prescient? A Continuous Evaluation using Daily News as the Oracle

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.652552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:61e60cac0bb2b4ddc58963d01b27c680b86d466487da9ee7718124fc7b63d91a

Observation b57a3ec0-f460-4c19-9234-395f07887415 · outbound

This paper cites an unresolved cited work.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:2b0dd438e15a7b29bba392351e526df6c164a593e0dcc7cc3e5d66887ede0e43

Observation fced5b50-938c-42ca-bccb-7ecd02fe663c · outbound

This paper cites Large Language Models Are Zero-Shot Time Series Forecasters.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Large Language Models Are Zero-Shot Time Series Forecasters

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.658392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:a299871d97dcc2c174cafe3aa4174367538074e3dbbc6dca55cfa8ba40ab03b2

Observation 444372f9-e854-470e-a635-d21d90470b77 · outbound

This paper cites The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.663170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:57fefe95d825d7458cea481aea2538ecf1b68db7ff9cff603078e94932919df5

Observation 81780c1f-5a6b-46b8-bfa6-f38d468a1e6c · outbound

This paper cites Time-LLM: Time Series Forecasting by Reprogramming Large Language Models.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Time-LLM: Time Series Forecasting by Reprogramming Large Language Models

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T10:58:02.659640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:896901246642cf6a339d2df85b783fbd91ffcdc59e244a3f29eb3fb7dfd839ee

Observation 5e4ac957-8140-47ce-8423-d5fb8e3f1cda · outbound

This paper cites an unresolved cited work.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:2ec178a948da7ae5e12be3ea4f0000d6ec60e4e061e6097544a42d873284ba64

Observation efb55a8c-9c5d-458d-9ae6-b5e451217b4b · outbound

This paper cites RealTime QA: What's the Answer Right Now?.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning RealTime QA: What's the Answer Right Now?

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.647009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:a90acdd2afe3f02469cbd7568589159094e6c500090652601ee2a0ba3e5694f4

Observation c5662c80-1290-4f54-962f-e887e3e5c8d2 · outbound

This paper cites Ramnani, Mayuresh Anand, Shubhashis Sengupta, and Andrew E.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Ramnani, Mayuresh Anand, Shubhashis Sengupta, and Andrew E

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-27T09:47:04.122464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:4c3c1cd18a3abb78263f5bce3e11f30aeb4c09885103102b72cf046904311ba6

Observation dc88334a-bc36-48a0-a189-80e0e38ceee6 · outbound

This paper cites Time Series Forecasting as Reasoning: A Slow-Thinking Approach with Reinforced LLMs.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Time Series Forecasting as Reasoning: A Slow-Thinking Approach with Reinforced LLMs

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-03T10:58:02.662179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:d5e06afec48602c2f2fb4ebaa888b24f453dddc7fa9b73aa64c0db82672f954a

Observation 2c83b60f-2bbf-4527-a194-a8e3e91e98b8 · outbound

This paper cites Proceedings of the 2023.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Proceedings of the 2023

Reference 36

Resolution
verified exact
doi, observed 2026-06-27T09:50:48.142123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:50c78f114292318537475509e3c05e52dea7984f0b6293588a7bc5187d9fc2b2

Observation 86111039-4e4a-4cb7-8d4d-0d291f5f31b6 · outbound

This paper cites Measuring Attribution in Natural Language Generation Models.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Measuring Attribution in Natural Language Generation Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.657429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:648d0a17e52051686e825ee74d3fb5961f45906fb518d152747c7d6b492f9105

Observation 257e21ac-37e8-4070-b7f8-53d640d9d2c0 · outbound

This paper cites an unresolved cited work.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Unresolved cited work

Reference 38

Resolution
verified exact
doi, observed 2026-06-27T09:50:48.139404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:39dea8a21e3acb58c32290f0437da06b12ce4c10faf2ac41999bf537e8203bc7

Observation 07a8b136-96c9-4270-8900-ea9e2f540c10 · outbound

This paper cites Bench to the Future: A Pastcasting Benchmark for Forecasting Agents.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Bench to the Future: A Pastcasting Benchmark for Forecasting Agents

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:58:02.649126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:823ebba4e08aac242120b39563860fe358d940c5a3fddd5478dbf00840885114

Observation ea1dc406-87a8-44cf-86d9-c7439340b30b · outbound

This paper cites ECHo: A Visio-Linguistic Dataset for Event Causality Inference via Human-Centric Reasoning.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning ECHo: A Visio-Linguistic Dataset for Event Causality Inference via Human-Centric Reasoning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.639448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:8429fe21a9de502899f63cbb1502a27d6810c7164175b44a0e7002c9eea7e41f

Observation 6adee730-89f9-449e-83cc-eac9dafb664b · outbound

This paper cites Cheng et al.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Cheng et al

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.641873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:bc48683e702acb7a391f569f482cd484e3a8e7fbf139049860c07fa2ee12026e

Observation 56df8d54-f063-4907-b5ff-78a3e3137172 · outbound

This paper cites MIRAI: Evaluating LLM Agents for Event Forecasting.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning MIRAI: Evaluating LLM Agents for Event Forecasting

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.665614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:e95e4d8538d31dd571373dcc0ceafd7ce31ab6ea2572b88bcc65c8389cb15f6b

Observation efeb700d-2193-4c1c-9936-07b826cc6462 · outbound

This paper cites InProceedings of the ACM Web Conference 2024(Singapore, Singapore)(WWW ’24).

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning InProceedings of the ACM Web Conference 2024(Singapore, Singapore)(WWW ’24)

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T09:50:48.134721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:e4f192839f61d3cb384811475e61e9c3745292d57997784d4ba0929566d9f2e0

Observation 974182e4-13b6-4d6f-947e-e8bd9afb1c85 · outbound

This paper cites FOReCAst: The Future Outcome Reasoning and Confidence Assessment Benchmark.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning FOReCAst: The Future Outcome Reasoning and Confidence Assessment Benchmark

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.660619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:bac14e58c94e24ab00410823f04ea8a4a58b1f72a70a9baa4ef8d56b2f6b935d

Observation 8c22fc98-30b1-46c3-bb82-0facb0491dc9 · outbound

This paper cites FutureX: An Advanced Live Benchmark for LLM Agents in Future Prediction.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning FutureX: An Advanced Live Benchmark for LLM Agents in Future Prediction

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.668078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:ac90c9a0310fa36d193e559d66e207ebbf996d8f2706812fdd2a3cd75bf96e03

Observation 7428a9ae-35b1-464b-b8ef-30d2dfd2f441 · outbound

This paper cites Is Your LLM Outdated? A Deep Look at Temporal Generalization.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Is Your LLM Outdated? A Deep Look at Temporal Generalization

Reference 46

Resolution
verified exact
doi, observed 2026-06-27T09:50:48.140762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:be9ecd5525728ca70e8668776843bb4bc81c575e9f3c467f5de27f8d212ba39f

Observation d229ea84-9cf2-4a51-a485-c325dcb1b1bd · outbound

This paper cites Forecasting Future World Events with Neural Networks.

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning Forecasting Future World Events with Neural Networks

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.671142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:47:04.122464Z digest=sha256:98a50fb33f3c41068ad275e173a92ab95e8683f10d67452e28ee56c5f35f1063

Pith citing papers

No inbound Pith citation observations are available.