Pith. sign in

Paper Citation Record · LEDGER

Generating Symbolic World Models via Test-time Scaling of Large Language Models

As of 15 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 3 inbound Pith citation observations for arXiv:2502.04728.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04728 v2

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:45:45.538720Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:14:21.027074Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T23:42:49.549067Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 03cd33d5-726d-41df-939a-57dd9a284dfe · outbound

This paper cites GPT-4 Technical Report.

Generating Symbolic World Models via Test-time Scaling of Large Language Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.239493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.239493Z digest=sha256:d47d89fb77774fafb9a19773ab686c539c7c5d5129c6d9af2d198a2439007035

Observation 0333c392-25f1-4a9e-baa1-91417723d13c · outbound

This paper cites Learning discrete world models for heuristic search.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Learning discrete world models for heuristic search

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.845605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.245571Z digest=sha256:9bc0b6b51ff2ccd894e1ee3413ca09fb47aa8655045d3372f5d3493b1c429ef7

Observation 36c5cbae-ff49-461e-acf8-180ea3d9df5d · outbound

This paper cites Program Synthesis with Large Language Models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Program Synthesis with Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.250220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.250220Z digest=sha256:7715db3ab76ff55ec93046deed032114a2cc43c5bae902bf77cf3d113d7f9da5

Observation 6d31b2ee-9447-45c1-8cab-ecccf8d49965 · outbound

This paper cites Learning warm-start points for ac optimal power flow.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Learning warm-start points for ac optimal power flow

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.831395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.255223Z digest=sha256:5ca00cf213d3e697eaa0efb7a3c28ca0b23f0d530ea605a7ecf798d639c1da6b

Observation 9e4b8858-2e3f-44c4-bf71-3ca19a79f864 · outbound

This paper cites Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.260053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.260053Z digest=sha256:8917315e221e17ee897877020ae163eec21a83ffeb7f9f07e4b2c8c2aec4cb30

Observation 39f13cc1-2b1d-46b8-a381-21be6515f46a · outbound

This paper cites OpenAI Gym.

Generating Symbolic World Models via Test-time Scaling of Large Language Models OpenAI Gym

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.265018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.265018Z digest=sha256:cae366e08089c4def7c177bb521f48fdc72c9726677e2616a284a2b2f0c4757c

Observation 85844672-b9dc-4afe-982a-cc1b173a2e71 · outbound

This paper cites Language models are few-shot learners.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Language models are few-shot learners

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.816778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.270167Z digest=sha256:cb39b32a42ca024f94cb66e9db2de77e96492720f2a4f11d8c45b6ec2c5e0c18

Observation bcbf0079-ae87-45ce-989a-c8b37f8c7cad · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Evaluating Large Language Models Trained on Code

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.275330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.275330Z digest=sha256:444dc74aa6a0b6acb29aadde275379801f084bde8731bff7de388689942f387c

Observation ede1b7a6-2aa2-4a1f-8c27-e9012feacaf2 · outbound

This paper cites Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.280344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.280344Z digest=sha256:d4f8961f8c0c1800f6dda14a53ab1327ac98d187f2b6f2ec997589d1a9387069

Observation f07b44fe-b95c-485b-ad97-09d742351135 · outbound

This paper cites Inductive or Deductive? Rethinking the Fundamental Reasoning Abilities of LLMs.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Inductive or Deductive? Rethinking the Fundamental Reasoning Abilities of LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.285649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.285649Z digest=sha256:6602fcef4ad1e5ba81e520b4bf965fa43249388730ad096b085f51cbea271f2c

Observation 0ea578a5-37f1-4a1d-a3e8-dc8c91c33b24 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Training Verifiers to Solve Math Word Problems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.290407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.290407Z digest=sha256:4ce16f7fc398943dc80e3ca115c447137f2363d051f4a773f517a219486c9fca

Observation e46c3c3c-c61a-4dc6-93bf-14d15caae9ad · outbound

This paper cites Generating Code World Models with Large Language Models Guided by Monte Carlo Tree Search.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Generating Code World Models with Large Language Models Guided by Monte Carlo Tree Search

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.295334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.295334Z digest=sha256:2aef10789762722527a1fe6912cf336f87833e2b0439603fa8d63c34154c6c26

Observation 9cae3b3f-2bc0-4940-a254-af6b3f0c1477 · outbound

This paper cites Parameter-efficient fine-tuning of large-scale pre-trained language models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Parameter-efficient fine-tuning of large-scale pre-trained language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.802075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.299976Z digest=sha256:4b8550300065cddfd1c1697858d399a67b118e38d3d26cf2a88d1df8c10098b6

Observation bba694be-40d2-4597-b281-d6988b038b87 · outbound

This paper cites A Survey on In-context Learning.

Generating Symbolic World Models via Test-time Scaling of Large Language Models A Survey on In-context Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.304298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.304298Z digest=sha256:6355955cb6efb387f4b6bac950bdaac13c01bf9e79233e850246ca776c83fce0

Observation eefa8b78-3ad8-45ac-92c2-2fa09093e5a1 · outbound

This paper cites The Llama 3 Herd of Models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.309015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.309015Z digest=sha256:ce2ec8a7f708c2941802ac6c7960ec0d0eca05f8f448aff1c1416c90a0551491

Observation 05e9d0c6-f8f0-4c7b-b83d-8b830a7848a5 · outbound

This paper cites Fikes and Nils J.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Fikes and Nils J

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.787696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.313744Z digest=sha256:55e1a9b5357c7ec12449171d7a98d5dc6bf849617bab316d723346b5f3e46327

Observation a8196531-c14f-4dd2-9fa5-80c78856f6b5 · outbound

This paper cites Leveraging pre- trained large language models to construct and utilize world models for model-based task planning.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Leveraging pre- trained large language models to construct and utilize world models for model-based task planning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.773058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.318349Z digest=sha256:681b10f3f8b5b527fb6fc4621bf0965c8f1cb7b4591969b9545799b86f030d14

Observation b7d8ce2b-c5dc-43e9-bd3b-a8549d7d7275 · outbound

This paper cites Reasoning with Language Model is Planning with World Model.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Reasoning with Language Model is Planning with World Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.322784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.322784Z digest=sha256:ff3e1d8d7b95e855ad3b7a8b9f03313b451d70d119bc7a3a3faae449e82369fd

Observation 7f0a7224-bc1f-4018-8a5d-da4c5e801049 · outbound

This paper cites The fast downward planning system.

Generating Symbolic World Models via Test-time Scaling of Large Language Models The fast downward planning system

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.758668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.327277Z digest=sha256:1a6bc26edfbcce4c9fd4e1b18da4a806c0abdf50a7d37d0ea6071cfd341b3b55

Observation ae4c5c4b-61de-4c3b-8da5-288a88e88961 · outbound

This paper cites The Competition: Impact, Organization, Evaluation, Benchmarks.

Generating Symbolic World Models via Test-time Scaling of Large Language Models The Competition: Impact, Organization, Evaluation, Benchmarks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.744061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.331678Z digest=sha256:283bbcde406ff27711c1581e62b142c2ede823013d3fca5ec5c78095543256c5

Observation 8f457a45-904f-4767-b2c9-6c79ece777c4 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Lora: Low-rank adaptation of large language models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.729161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.335951Z digest=sha256:c0b2e0b91ebc283ab1f7e33b3017f3c55fd20e8d729008eb62f25ae89b13e2b3

Observation 6a196f09-0952-4697-b545-3cb315e96399 · outbound

This paper cites OpenAI o1 System Card.

Generating Symbolic World Models via Test-time Scaling of Large Language Models OpenAI o1 System Card

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.340311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.340311Z digest=sha256:e3e2df3b77be4d740ee3cca6563ed1c3c27e9d2f0486913992c64aab997343ff

Observation 646869af-157b-450d-b043-6057ffe4fe10 · outbound

This paper cites Mistral 7B.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Mistral 7B

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.345137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.345137Z digest=sha256:bcfeea637dba29416d1526fa3cfa0e17775aeb7f6908ea3e78df87fea43f0304

Observation 1a73d1ce-9422-42c9-910b-472c4d627b15 · outbound

This paper cites Can large language models reason and plan?Annals of the New York Academy of Sciences, 1534(1):15–18, 2024.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Can large language models reason and plan?Annals of the New York Academy of Sciences, 1534(1):15–18, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.714844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.350020Z digest=sha256:276a28972545ac4f76eea43c6c395fd65164206ec64566cd83be3c4205a5fe1d

Observation f93ed7bc-4342-4b77-8817-7936a8abda61 · outbound

This paper cites Parameter-efficient orthogonal finetuning via butterfly factorization.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Parameter-efficient orthogonal finetuning via butterfly factorization

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.699971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.354281Z digest=sha256:b45496f1445e6ed61a9b366e1bf08f8da648ea1d04da30ed9357143249893a05

Observation 946a65a8-9d97-4a61-a89a-5c4701501fea · outbound

This paper cites Leveraging Environment Interaction for Automated PDDL Translation and Planning with Large Language Models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Leveraging Environment Interaction for Automated PDDL Translation and Planning with Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.359163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.359163Z digest=sha256:04ddcf5188fb0050bf646c0788826d18d6ae2f54904c34cb8a7f561f646a1a34

Observation 54cfd0fa-f274-448c-b515-e2a85c50e5a1 · outbound

This paper cites Howe, Craig A.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Howe, Craig A

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.685254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.363988Z digest=sha256:83514e588b79752393e042a028a7e6eda9476d81f5f8c01106709506a43cf71a

Observation 45ab39e7-1ebf-4626-884c-998a24d671d0 · outbound

This paper cites GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.368651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.368651Z digest=sha256:8f669f0ac66402f48cb83b319191f40d7f591cb1f1f32c3311c72d59a7cdbe82

Observation b772d60b-b75b-43d3-831e-f8a9906aac37 · outbound

This paper cites Fully autonomous ai agents should not be developed.arXiv preprint arXiv:2502.02649, 2025.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Fully autonomous ai agents should not be developed.arXiv preprint arXiv:2502.02649, 2025

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.374118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.374118Z digest=sha256:01bcff899468c8a22f53f4e083df4d26b5fb5d0404ea90385804b6b4dfc9e402

Observation 68346e29-ac90-4df6-99d3-75a7b5ab64fc · outbound

This paper cites Large language models as planning domain generators.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Large language models as planning domain generators

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.671010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.378530Z digest=sha256:071e9b9e83a0a0aab1c7aef4008fc7b87a6047adf8ee96daad5557bb39b6148f

Observation ed5725ed-d274-40e8-89c1-761ee62b7939 · outbound

This paper cites Automatic Prompt Optimization with "Gradient Descent" and Beam Search.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Automatic Prompt Optimization with "Gradient Descent" and Beam Search

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.383025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.383025Z digest=sha256:5499633f53603460edfe8d6fb31f78e2820332299813eae62cc8cba908ee3cdd

Observation 654adc1e-7ed5-4d85-901d-6daa106d887e · outbound

This paper cites O1 Replication Journey: A Strategic Progress Report -- Part 1.

Generating Symbolic World Models via Test-time Scaling of Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.387659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.387659Z digest=sha256:fca65db58f17acb887ea67755c3a56aab02a146a0e89484a93d05c831525b19b

Observation d184fcdf-b6bd-47af-9930-beae695c542e · outbound

This paper cites Controlling text-to-image diffusion by orthogonal finetuning.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Controlling text-to-image diffusion by orthogonal finetuning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.656622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.391868Z digest=sha256:6dc4e9b4cf3a3ba8d358f7c4675354a615a88847b020e94bed172c1842e3ce5b

Observation 0713ec78-5d08-4f1c-bdff-b5e5fc75dd4e · outbound

This paper cites Pearson, 2016.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Pearson, 2016

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.637822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.396629Z digest=sha256:f3542ace00bcc4f058b064a2a71a5cc5cf357c1d1e9a6475814e832836cb9a42

Observation 9fadc9f0-a05a-4a15-8305-74c25f31662e · outbound

This paper cites Learning Multiple Initial Solutions to Optimization Problems.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Learning Multiple Initial Solutions to Optimization Problems

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.401048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.401048Z digest=sha256:c236096b9c526f10f73e795b20298b0de30f26013ea5a06a39017edaced33e14

Observation 57bc6fa8-88cf-4830-a177-cab53de0f5b8 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.Advances in Neural Information Processing Sys- tems, 36:8634–8652, 2023.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Reflexion: Language agents with verbal reinforcement learning.Advances in Neural Information Processing Sys- tems, 36:8634–8652, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.622404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.405565Z digest=sha256:e7cb98bcc2ac5dc4b56f3f91645d9f84b00471bd26bbf4ca053018eec7295cc2

Observation 7a02912c-cd62-4dad-b69b-5270c4090adf · outbound

This paper cites ALFWorld: Aligning Text and Embodied Environments for Interactive Learning.

Generating Symbolic World Models via Test-time Scaling of Large Language Models ALFWorld: Aligning Text and Embodied Environments for Interactive Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.409955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.409955Z digest=sha256:da408312f048288f003ee53916d7548886130599e8026702941aee62a432dcaa

Observation aa6afb8f-1814-45f2-a363-8dae05abc82a · outbound

This paper cites Generating consistent PDDL domains with Large Language Models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Generating consistent PDDL domains with Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.414679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.414679Z digest=sha256:eb6f67488f6b6f1c8f3b1853a2dd6fc1c8d2e32ab5c7993ffd9c0c1de434b38b

Observation 8b811352-2a25-4087-9114-a9c5b07014ac · outbound

This paper cites Inference scaling flaws: The limits of llm resampling with imperfect verifiers.arXiv preprint arXiv:2411.17501, 2024.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Inference scaling flaws: The limits of llm resampling with imperfect verifiers.arXiv preprint arXiv:2411.17501, 2024

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.419311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.419311Z digest=sha256:312ab3c5f59935e004ee9ff2cf49f4f6af9dbc69e54a2bf01662f4f353e9dd7c

Observation 138c330b-e603-41c1-9383-3aebb683c63d · outbound

This paper cites WorldCoder, a Model-Based LLM Agent: Building World Models by Writing Code and Interacting with the Environment.

Generating Symbolic World Models via Test-time Scaling of Large Language Models WorldCoder, a Model-Based LLM Agent: Building World Models by Writing Code and Interacting with the Environment

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.423844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.423844Z digest=sha256:8dd11e221b19dae699fe7a89c55a92c281167cbe46a64bf85367acff7ab3718e

Observation 02887bf2-501d-4228-854e-948372ef04ff · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Gemma: Open Models Based on Gemini Research and Technology

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.428515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.428515Z digest=sha256:30f30a65d35bd0bac2584e78b6ea5e527bff1bea1f164001790dbb770e4fdf4f

Observation 320494a9-972c-4a9b-aaf3-4c3a54ad206c · outbound

This paper cites Planbench: An extensible benchmark for evaluating large language models on planning and reasoning about change.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Planbench: An extensible benchmark for evaluating large language models on planning and reasoning about change

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.607314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.433859Z digest=sha256:c5bc0d2b99fa7474608b273eee37442544f831037e486f45ded3040fec28848d

Observation d1ec938e-c08e-4c54-8447-311a147ecf4c · outbound

This paper cites LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench.

Generating Symbolic World Models via Test-time Scaling of Large Language Models LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.438262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.438262Z digest=sha256:8247e3cfa0f0e72858771453b779cf5812fa0cad0943b793ce54754e3dc8fc99

Observation 03ae988b-5ada-4c24-970e-d209af5d3491 · outbound

This paper cites On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability.

Generating Symbolic World Models via Test-time Scaling of Large Language Models On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.447954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.447954Z digest=sha256:cfaaccef23d6816aee55fd18dc5650d1be9b369d74a9ced7bd52571e9dcb79eb

Observation 2b2e7a1c-ce55-4a01-8229-9cd14def1877 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.452286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.452286Z digest=sha256:a4849602cf18f4f9075d23d303e6dd3b9b18b242993237d1ae3335a56411c172

Observation 26aedf24-c089-4a68-a1c2-24a2e02fbdb8 · outbound

This paper cites Emergent Abilities of Large Language Models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Emergent Abilities of Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.457229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.457229Z digest=sha256:919bd4a2c3216659b46204eaab1a361b35d3961939878eb049258c70f9a78c29

Observation 7e4dc49f-163a-48a8-ad00-e8ea8df8ae46 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Chain-of-thought prompting elicits reasoning in large language models

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.592486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.461751Z digest=sha256:8e11c8f649c6df57cc21b6a990078319b5f082b2df767dc3ac4bacb1b50ce969

Observation b7151f06-4a35-470d-95bf-a263b660de2d · outbound

This paper cites System 2 Attention (is something you might need too).

Generating Symbolic World Models via Test-time Scaling of Large Language Models System 2 Attention (is something you might need too)

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.466137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.466137Z digest=sha256:794b38aadcf4586ed1368b085ba8db8697e30f3f168ae60da4bf587de40a8399

Observation 1045e54b-a72d-44d0-a6ec-35c66fa8dbae · outbound

This paper cites Verbalized Machine Learning: Revisiting Machine Learning with Language Models.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Verbalized Machine Learning: Revisiting Machine Learning with Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.470962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.470962Z digest=sha256:93c3a44b4faaed9cd90f0aa04f5821fcb00c5d6beace73b725c60557854a3bc3

Observation 670baf85-dec8-40a9-8b0a-1045fe5539a2 · outbound

This paper cites Qwen2 Technical Report.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Qwen2 Technical Report

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.475667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.475667Z digest=sha256:c97f7d2f8d58adbb00f871e8391c05640368a72ce082628f853a57a009f3ca2d

Observation 0db9d5d2-2838-4684-adc7-5b4f1e931cde · outbound

This paper cites Le, Denny Zhou, and Xinyun Chen.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Le, Denny Zhou, and Xinyun Chen

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.576067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.480493Z digest=sha256:ad962b158646139df66b91c487b167926a3631780a8004a6146c4122bd9a4ca0

Observation 5d3b7035-f87b-412a-b727-b261da42ae26 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Yi: Open Foundation Models by 01.AI

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.484908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.484908Z digest=sha256:feadf37e6be213e52c74bf46f166c76e380a2bd1756ebb1408eb7c9a8b0c114c

Observation 41055a6c-f42a-4147-ae5f-31763f762d3b · outbound

This paper cites Distilling System 2 into System 1.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Distilling System 2 into System 1

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.489671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.489671Z digest=sha256:2e3949e6441ba34190054d08c357ed8a573f3fb366576556b14f2bfa916a7f45

Observation cb0d4312-531b-41b3-a5bd-24554ff9ce11 · outbound

This paper cites TextGrad: Automatic "Differentiation" via Text.

Generating Symbolic World Models via Test-time Scaling of Large Language Models TextGrad: Automatic "Differentiation" via Text

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.494452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.494452Z digest=sha256:7f5550e04ae4e66fa46f625335aec566a16a92b97ad1f41b1365c7b52665a4ba

Observation e390a31d-ece3-4680-8602-4e9b7f4b24a3 · outbound

This paper cites Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.499263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.499263Z digest=sha256:2594f8d48119b5e5426620c784b3e92c322696ff5d16860d66d4842a6c027fce

Observation 0135e484-0cab-4eed-8cc7-889552487b0a · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.504109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.504109Z digest=sha256:94498627eacb563c7e5618d88404ef138e4703153951b55b134148474885ee83

Observation e9755324-e6a7-4746-9e07-35f758271294 · outbound

This paper cites MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics.

Generating Symbolic World Models via Test-time Scaling of Large Language Models MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.509398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.509398Z digest=sha256:636b58391732ef2c834a6a2c0746cc74f63bab28081899a34cbd4c55896fe29b

Observation f78d30ca-a71c-44cf-bfd8-29d54df6cd81 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.561071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.515011Z digest=sha256:62044a95f3a64ed5a23efead3de0999e96d3dab93334698bd9362f3985c1d418

Observation ac2a8a49-3c8e-4897-92cb-edec6482a32d · outbound

This paper cites Large Language Models Are Human-Level Prompt Engineers.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Large Language Models Are Human-Level Prompt Engineers

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.519535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.519535Z digest=sha256:a9a3c5cb296e87b6d1838a85cedef9ee0fcb64ee2e5e0eaf60a7c6db5035145d

Observation 0017a679-4b3c-4983-9544-091018a6b5b0 · outbound

This paper cites Plane- tarium: A rigorous benchmark for translating text to structured planning languages.arXiv preprint arXiv:2407.03321, 2024.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Plane- tarium: A rigorous benchmark for translating text to structured planning languages.arXiv preprint arXiv:2407.03321, 2024

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.524359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.524359Z digest=sha256:e6172ae4c54df019aaf1cb05139ed72ce39f2b0049fc6bd4eceea80096c0a92b

Observation 87adb92d-79da-48f4-b24a-c65f2c4e2186 · outbound

This paper cites object1 is washed and heated.

Generating Symbolic World Models via Test-time Scaling of Large Language Models object1 is washed and heated

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T21:45:46.545188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.529282Z digest=sha256:06b4da7b07ea912c01888c1705eb7581059496ffaa181a2ac13bbfd70b5d03b6

Observation f993925a-28f3-40d8-bf4f-44d289741275 · outbound

This paper cites an unresolved cited work.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-08T21:45:46.528926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.534417Z digest=sha256:278e9ac5e4b0d59c84e92a62de5ccf94752a1809b597b9f95b8a5e161085457c

Observation 6022571f-e1b0-4246-8129-db0ae523b999 · outbound

This paper cites an unresolved cited work.

Generating Symbolic World Models via Test-time Scaling of Large Language Models Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-08T21:45:46.513526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T21:45:45.538720Z digest=sha256:1189dd035fd2df4be8a38216041763952cf9d3eb50c07d6f527955b280ad3029

Pith citing papers

Observation d1d36373-ba35-402f-b684-c848f9aea5ad · inbound

CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization cites this paper.

CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization Generating Symbolic World Models via Test-time Scaling of Large Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:21.027074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:14:21.027074Z digest=sha256:ecfd63f4b704b9efb4f430c51f0ba4920535002b9932a84fae3345b72f0a1cc8

Observation fdba0c82-df4a-4742-a488-ce638e579571 · inbound

Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks cites this paper.

Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks Generating Symbolic World Models via Test-time Scaling of Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T06:05:30.314068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:05:30.314068Z digest=sha256:14c2a5f7cd77745a1b97e1b5ae747385478e0efca0e72b9c0a90c1600c2084b6

Observation b55257f8-e899-462e-a68e-8fb55168961d · inbound

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning cites this paper.

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning Generating Symbolic World Models via Test-time Scaling of Large Language Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:42:49.550411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T23:38:38.127345Z digest=sha256:d72bb888b8506370bb437bd5a29af96e3e855ac52bbe4e006fec6a7cff391996