Pith. sign in

Paper Citation Record · LEDGER

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search

As of 7 August 2026, this Paper Citation Record lists 100 of 217 outbound references and 1 inbound Pith citation observation for arXiv:2507.00004.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00004 v2

Coverage vector

measured 100 of 217 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:07:39.908059Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-02T12:29:24.439779Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T12:36:56.135904Z

Reference resolution

100 of 217 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved100
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ded0eeef-24b1-45e1-be13-4bc79ceb2429 · outbound

This paper cites Scaling Laws for Neural Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Neural Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.492720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.492720Z digest=sha256:e53d3fe00bd36e17005a734a47fceda367c29cf119f47d4115e9f9b3d26a6c42

Observation bb086a52-8788-4114-a846-700cf8dbfb0e · outbound

This paper cites Training Compute-Optimal Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training Compute-Optimal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.498613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.498613Z digest=sha256:c8378b0a6b98ca3b088dbc8d9446dead19a1d3fd0fdbb301204fbe63ada68a5e

Observation b0430e10-3c59-4534-997b-f8613924f1d5 · outbound

This paper cites Compute-Optimal LLMs Provably Generalize Better With Scale.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Compute-Optimal LLMs Provably Generalize Better With Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.503440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.503440Z digest=sha256:d2d77c7730cb5d394cb141fa8c07c5c4fef601f2b984cb4d1a7113f616dad1f8

Observation f0d827f6-3399-4524-9224-f3b73361d949 · outbound

This paper cites Training compute of frontier AI models grows by 4-5x per year,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training compute of frontier AI models grows by 4-5x per year,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.507915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.507915Z digest=sha256:814ba662fc66efc9b4fbd4877a3eca6291a5c99baf92f89eb16083fd2a525419

Observation 5efef214-44c5-4040-8b6f-e7af7d39138a · outbound

This paper cites Increased Compute Efficiency and the Diffusion of AI Capabilities.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Increased Compute Efficiency and the Diffusion of AI Capabilities

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.512826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.512826Z digest=sha256:b6e745a37cb8699eacc73daf74a931484748a39d793cb26a10c386b56bc3ea7c

Observation 983dcfe7-b288-45fb-9d9d-5b300cfdfd01 · outbound

This paper cites Measuring the Algorithmic Efficiency of Neural Networks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Measuring the Algorithmic Efficiency of Neural Networks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.517278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.517278Z digest=sha256:02a7cb871dd44145c9d6482a2caefbbde759108caa01294dd40a89d89c0bedde

Observation a40e8546-a6c2-4358-9f2c-4b0cc46c1a01 · outbound

This paper cites Algorithmic progress in language models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Algorithmic progress in language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.522279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.522279Z digest=sha256:b3c785596efb29db964041aea0d95f6406600c1375d49bc9a268a67e38be654f

Observation 55527bc7-82f6-45e1-a4e7-de716a995acb · outbound

This paper cites Claude’s extended thinking,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Claude’s extended thinking,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.526454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.526454Z digest=sha256:478c85b568a2f7fa6eec771e415e89c8347d2d8b4ef29010ddcc0c33e46fb82f

Observation 94aeaad4-b318-4e1c-8996-82fa3d566d7b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.530848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.530848Z digest=sha256:588aae82b9432a567fc973ea3702834706a5850722bd21c38cb72c57a80e00c1

Observation d176cbbe-b1c1-4eeb-9c9f-b04f7c48f3cd · outbound

This paper cites Gemini 2.5: Our most intelligent AI model,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Gemini 2.5: Our most intelligent AI model,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.534878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.534878Z digest=sha256:e61a88ea7940bb1c5f947401322ca854e2d6b2965b90c362dc6d9605d9513641

Observation d7f1571f-5cf1-4269-b93e-c61995d18531 · outbound

This paper cites IBM Granite 3.2: Reasoning, vision, forecasting and more,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search IBM Granite 3.2: Reasoning, vision, forecasting and more,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.538768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.538768Z digest=sha256:8f50acb57d09ca2972bdc5c1c4224eec1dcdfaca7ec210f60b470fa8e04c8465

Observation be84ecfc-bee3-440e-87d9-358799f52e61 · outbound

This paper cites Phi-4-reasoning Technical Report.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Phi-4-reasoning Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.542544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.542544Z digest=sha256:1dc766f4ec60c8a2418fdc10c2b4b1bf2da99a200e5fbc459efce21d3c909ff7

Observation 1fc145bf-976b-4c12-b63e-7551b5072a6c · outbound

This paper cites OpenAI o1 System Card,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search OpenAI o1 System Card,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.546640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.546640Z digest=sha256:53e4b11c8dfcee5913873e40ce363969c9c62be7fbfa773fb6e7d1089c3e35f4

Observation 0201b25d-8625-4501-8250-aa7f9357c89b · outbound

This paper cites Introducing OpenAI o3 and o4-mini,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Introducing OpenAI o3 and o4-mini,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.550360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.550360Z digest=sha256:029816693cb8dfdbe90ea999cf424c6e9af0c533a7109da022bb10285012e4ba

Observation b1190000-cdb1-46d9-90f8-e9f8ffd15ccb · outbound

This paper cites Grok 3 Beta — The Age of Reasoning Agents,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Grok 3 Beta — The Age of Reasoning Agents,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.554243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.554243Z digest=sha256:8ee42becc0240e4f1d48a1fb327e8578f9ab87e1b4ada003607d230c9e07a8f1

Observation be22cfff-d3eb-4ee5-b11f-ad8d106bbf3b · outbound

This paper cites The growing energy footprint of artificial intelligence,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The growing energy footprint of artificial intelligence,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.557955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.557955Z digest=sha256:c9106d42d1998aa3878223403e7e9ce60248b8ea4c58c89f3f109e2174f267af

Observation 77f11932-1a2d-439a-9e21-ecd23f414404 · outbound

This paper cites Estimating the carbon footprint of BLOOM, a 176B parameter language model,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Estimating the carbon footprint of BLOOM, a 176B parameter language model,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.562176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.562176Z digest=sha256:30897b13d6a1bb7f1be6247c83c323f0218eaef30b0a3535a5042860ced0ae25

Observation a8743a1b-2298-4a01-9984-c89d0c621898 · outbound

This paper cites The Carbon Footprint of Machine Learning Training Will Plateau, Then Shrink.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The Carbon Footprint of Machine Learning Training Will Plateau, Then Shrink

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.565848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.565848Z digest=sha256:6340d9f5b2f3ce6b300d00039ea7e04675c73690a37abc1688f149cb5e642f25

Observation d5477b62-9f63-4959-9c94-04557d63739a · outbound

This paper cites Sustainable AI: Environmental implications, challenges and opportunities,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Sustainable AI: Environmental implications, challenges and opportunities,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.570414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.570414Z digest=sha256:38d9b6d6274b2f6bff64ba21d053906922b4fb003304e07207e77296a23f9366

Observation 43c6e430-13ee-4da6-b7ea-0fefbd694ff9 · outbound

This paper cites The next wave of AI: Demand and adoption,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The next wave of AI: Demand and adoption,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.574308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.574308Z digest=sha256:6a406819b4ee3b44a82b361dd57a175d32d4d4666629644d9eef80383b8cb3d0

Observation 96a5ee41-d563-4a26-8ec6-1a7235947640 · outbound

This paper cites From Efficiency Gains to Rebound Effects: The Problem of Jevons' Paradox in AI's Polarized Environmental Debate.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search From Efficiency Gains to Rebound Effects: The Problem of Jevons' Paradox in AI's Polarized Environmental Debate

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.578208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.578208Z digest=sha256:36d76b840a60a1247fdf4efeda7ced3431ca7832d4da8453b72d5592f796dc44

Observation bd8f6d77-36a2-4910-a6a6-dcf1bfd64aae · outbound

This paper cites an unresolved cited work.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.582584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.582584Z digest=sha256:154cf71807b2c11a68127e713e17f8e21c07d8eaf697d8d94ad050634e15b903

Observation 4cb6bb14-8ced-428d-8814-8dfb901fea59 · outbound

This paper cites 1B user messages sent on ChatGPT every day,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search 1B user messages sent on ChatGPT every day,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.586504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.586504Z digest=sha256:bcbf5a6d01114b9b2980c73a9137327ab3174584e1e761f9392060f87c5eb395

Observation b3a8d728-3b7d-46d4-879a-679fc2647aee · outbound

This paper cites ChatGPT added one million users in the last hour,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search ChatGPT added one million users in the last hour,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.590585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.590585Z digest=sha256:32c401f68a8bc39b77e0c99e7dafe8dc8d4766423ca3a260adfe29edf473cf09

Observation 97a61648-86eb-4065-9d7d-d6e0b30c8fbf · outbound

This paper cites ChatGPT statistics and user trends (2025),.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search ChatGPT statistics and user trends (2025),

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.594451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.594451Z digest=sha256:1e6cb6c2ff236f64ba1b86f293d99c86a131a956bb52a711c1c811a28cba4e6f

Observation c6daaac8-50e7-423f-b605-e0e4e0a7907d · outbound

This paper cites A systematic review of Green AI,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A systematic review of Green AI,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.598332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.598332Z digest=sha256:eac453749eb7860c1fe62f02c1eb9c0c1691f9a60f65c226b505f1d8409b487e

Observation 80a688e5-b70a-40ea-91c4-956cb869db40 · outbound

This paper cites Deep Blue,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Deep Blue,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.602174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.602174Z digest=sha256:a3f38d705e225fe1ab3173ca9c5b745cd00963c73911510f95ef60640285e3a0

Observation 6e7aebae-75b6-4a46-b0a6-9400c32231a5 · outbound

This paper cites Mastering the game of Go with deep neural networks and tree search,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering the game of Go with deep neural networks and tree search,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.606172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.606172Z digest=sha256:4914d24da478c123f007255c7ca1c43f9e875c0c5b2627f4c98cc7540b9c830a

Observation e5c0c3f8-e05a-491c-b108-54348491f428 · outbound

This paper cites Mastering the game of Go without human knowledge,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering the game of Go without human knowledge,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.611183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.611183Z digest=sha256:7d77a81af9f230c0d936335da230ff03316d38ad5e80f02ccd2a57c56acb824a

Observation 3d9dc32f-db41-44a7-8c75-c83962b8fee6 · outbound

This paper cites Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.615289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.615289Z digest=sha256:32c107aa3cf03f79475595da6f482ffb1ab2ec52fbbd12ba08384c48ecef764b

Observation 6f8b12d7-0413-43ee-87ec-b1f2fa7277c0 · outbound

This paper cites Scaling Scaling Laws with Board Games.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Scaling Laws with Board Games

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.619445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.619445Z digest=sha256:c686e1053dfa44779a86a1ce384e56274208ba53eac61bb5db6004228634e054

Observation 638d68f3-0448-4836-bc9a-b4bb22aa41b0 · outbound

This paper cites Safe and Nested Subgame Solving for Imperfect-Information Games.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Safe and Nested Subgame Solving for Imperfect-Information Games

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.624213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.624213Z digest=sha256:783d900b5bf2be1eb181cc2dc26b68988a54b9369a1dcf1c007a7f58553f6f3d

Observation a549f08b-343c-4022-aca5-e6358296ec98 · outbound

This paper cites Human-level play in the game of Diplomacy by combining language models with strategic reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Human-level play in the game of Diplomacy by combining language models with strategic reasoning,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.628464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.628464Z digest=sha256:f536e3418bf4dcdf642bdcc56891e934d40139638e2df9c6d0e94a9561bd65a0

Observation c111cbba-fc11-4128-a660-941e7adc9728 · outbound

This paper cites Emergent Abilities of Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Emergent Abilities of Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.632468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.632468Z digest=sha256:9b35b9a42dae53d249f7c6a87abdc3bc1ec90ddc5036f25a96b68bc7d32f9239

Observation f63715f7-46e5-4879-84ca-a789399dc89d · outbound

This paper cites An information theory of compute-optimal size scaling, emergence, and plateaus in language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search An information theory of compute-optimal size scaling, emergence, and plateaus in language models,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.636674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.636674Z digest=sha256:817d27affdf156f671f42fae38336ea58262a62b08bc1448c0ee31e64c7224cb

Observation d0251886-4a32-4390-8104-c926f87aa3af · outbound

This paper cites Multi-task Language Understanding on MMLU Leaderboard,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Multi-task Language Understanding on MMLU Leaderboard,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.640705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.640705Z digest=sha256:5a174664179bda6dabf985fe63d918e224c0f69478f82e15b0c8c82e7bee0b01

Observation 29c7dd85-4868-465f-8763-e1222da9e0d5 · outbound

This paper cites Measuring massive multitask language understanding,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Measuring massive multitask language understanding,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.644826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.644826Z digest=sha256:0a5e6a71290bb36e854df2623a2d41731e615cf815eb4544a23accf3d39c92b4

Observation 2c45d034-17db-427c-a26d-ef2b36bfdb98 · outbound

This paper cites Are emergent abilities of large language models a mirage?.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Are emergent abilities of large language models a mirage?

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.648721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.648721Z digest=sha256:31580daa36fef30ff88862aa261dbe84686c6322299114787a4a9eab37646f40

Observation e0d401d8-160a-41a9-9e0f-bb25ed697939 · outbound

This paper cites The quantization model of neural scaling,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The quantization model of neural scaling,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.652682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.652682Z digest=sha256:bef6f00b0e1c5fcca16dc9bc04a4826614d6291bec170b2109cddbf7ed88b98b

Observation e6616e3c-20cf-4f50-8c31-bc0dc045550a · outbound

This paper cites Circuit tracing: Revealing computational graphs in language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Circuit tracing: Revealing computational graphs in language models,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.656514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.656514Z digest=sha256:c33c8b19651505749f5f715c1364acef210828845b32a140533b6e9884e0d725

Observation 3fa94251-d92d-4778-9dc5-ae14f5e8ac3f · outbound

This paper cites Curriculum learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Curriculum learning,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.664621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.664621Z digest=sha256:04f2ca9c22733f47b44addcc408746b2d8b139a77e52b4fb2e526ed53e55552e

Observation 86c4006f-a309-4535-b9cf-08632f0485c9 · outbound

This paper cites A Theory for Emergence of Complex Skills in Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A Theory for Emergence of Complex Skills in Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.668555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.668555Z digest=sha256:41b88ba9007b8024ed28de133cb23fb3aee3bc1e17f78b3e6fea33845a4d565b

Observation c12e0140-4c77-4313-8279-37d316de2be2 · outbound

This paper cites A mathematical theory for learning semantic languages by abstract learners,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A mathematical theory for learning semantic languages by abstract learners,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.673016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.673016Z digest=sha256:e7ee4a2497bb3cfee170cdd635b2d6747e03df7e39d086b9b081810eddc54ee8

Observation 66468c21-ea12-4cdf-83f6-61997079e119 · outbound

This paper cites Skill-Mix: a flexible and expandable family of evaluations for AI models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Skill-Mix: a flexible and expandable family of evaluations for AI models,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.676978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.676978Z digest=sha256:62ae8685403d5f453957fc2044534a6fac1f92aa44ddb4280cb3a85613a521a8

Observation 0cd33ccb-1cbf-4e55-ad77-ba6e7bff0b80 · outbound

This paper cites The learning curve: implications of a quantitative analysis,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The learning curve: implications of a quantitative analysis,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.680836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.680836Z digest=sha256:b7cc1102a249c940baa648749d975daaca044871375ea0a3ad6ab0090de73708

Observation 27c29121-cbce-4bce-83b4-97c2df0c216f · outbound

This paper cites Plateaus, dips, and leaps: Where to look for inventions and discoveries during skilled performance,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Plateaus, dips, and leaps: Where to look for inventions and discoveries during skilled performance,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.684918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.684918Z digest=sha256:d31a806533df20891ada8c64b9e61e209d69610cf185e00642b21c1eb3396235

Observation 91c8e278-8ca1-4dfa-927a-b870f9af7cd7 · outbound

This paper cites A first-principles mathematical model integrates the disparate timescales of human learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A first-principles mathematical model integrates the disparate timescales of human learning,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.688862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.688862Z digest=sha256:6e4db3f4170271df2921a0a9604d8b322047b2a32944cce329840b442443936a

Observation 6bd50324-3c6f-4f18-971b-9b663a5a8c9f · outbound

This paper cites Spin-glass models as error-correcting codes,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Spin-glass models as error-correcting codes,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.693324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.693324Z digest=sha256:f061261ad6b0cf0d9f852435fe8289c2ecd20e271e9a677c58f396ed73a023d4

Observation 8491fe46-a0cd-40d9-9638-2b9ad7f199dd · outbound

This paper cites Newell,Unified Theories of Cognition.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Newell,Unified Theories of Cognition

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.697193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.697193Z digest=sha256:c127252205dc179da345bddb995509a067c01e601cd0252c91f1ac1fac643b59

Observation 4c225486-1177-463a-bbeb-f745387c52fb · outbound

This paper cites Barab ´asi,Network Science.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Barab ´asi,Network Science

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.701454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.701454Z digest=sha256:2b180a8f8afac6e7d9f7890757aed8629fb9ccc9ae406bde636eee5a342a4276

Observation b721141f-d0e6-4ec5-b9df-f670f1449ae9 · outbound

This paper cites Learning curves: Asymptotic values and rate of convergence,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Learning curves: Asymptotic values and rate of convergence,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.705522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.705522Z digest=sha256:1165bd02bfec45a03d7061841c704ae9286372af06636df0e608bb7418bc95fe

Observation 27dac46d-595c-4af0-be03-16caa56f7f4b · outbound

This paper cites Deep Learning Scaling is Predictable, Empirically.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Deep Learning Scaling is Predictable, Empirically

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.709672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.709672Z digest=sha256:65cdf273730e9c68523319194d5a99e63c37bceb773f95acf3aa97abebd7e856

Observation 77e77b35-f651-4a00-96d2-619bda196821 · outbound

This paper cites A Constructive Prediction of the Generalization Error Across Scales.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A Constructive Prediction of the Generalization Error Across Scales

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.713752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.713752Z digest=sha256:3e372e5ba3f28a92c2566d8ff597cc7067b11a4ed148a1d09b83f33d8f10f966

Observation b2483da6-1c56-4ecf-b3e0-8f16bcb95355 · outbound

This paper cites Language Models are Few-Shot Learners.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models are Few-Shot Learners

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.718254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.718254Z digest=sha256:6da5b4cb1f5db1742d8b1589bc43f6a5e0ff6f1e2b492f1eee449c5dcedf2e87

Observation a5a4d548-51d1-4c27-b5a3-208835a5ce7a · outbound

This paper cites Prediction and entropy of printed English,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Prediction and entropy of printed English,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.722424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.722424Z digest=sha256:adc8ccdc10b19f7f66fbb359455ff8d7636c8c1f2c3854aa91bf8a99ad524fda

Observation 6366e3eb-1eff-489f-9840-95594bbd5bf6 · outbound

This paper cites Explaining neural scaling laws,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Explaining neural scaling laws,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.726859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.726859Z digest=sha256:fc0909b98466ed2cefde71054090f73d2e6bbe6d8c0c6b594975a099c0b4f5df

Observation b54a65a2-cecf-44f0-8acc-597adfd2fe59 · outbound

This paper cites Towards a universal scaling law of LLM training and inference,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Towards a universal scaling law of LLM training and inference,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.730817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.730817Z digest=sha256:130a3b84fee10c7b2130f6de355b58aefc864560ceb574b9ddebb3bd316c7314

Observation 1e02fe25-2a3f-4352-8ea8-8dba24901087 · outbound

This paper cites The Llama 3 Herd of Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The Llama 3 Herd of Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.734911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.734911Z digest=sha256:c31b12278c847490f27a5035764d30b90b8eaaa880fd93a7b7316635a0a74da8

Observation 7786d76e-84fb-4402-a578-d68742c0b14c · outbound

This paper cites Densing Law of LLMs.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Densing Law of LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.739462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.739462Z digest=sha256:d9a3da39bf7c9187d9b404052cdd781551233271981d322b162af3ea2a84a85c

Observation 341e967e-f53c-44d0-b95a-640968e74a98 · outbound

This paper cites How predictable is language model benchmark performance?.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search How predictable is language model benchmark performance?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.743953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.743953Z digest=sha256:d91a78537a1812a2c54b29911e1ca8f5ec712a48ecc1692ea22fb2f79346ba71

Observation 55d90686-d264-4dcc-acd3-a88ee28203c9 · outbound

This paper cites Observational Scaling Laws and the Predictability of Language Model Performance.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Observational Scaling Laws and the Predictability of Language Model Performance

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.748185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.748185Z digest=sha256:fa9b0df68c9c29a34f914b5f002dbd7fa065cb8a9e8f29357fe7109be4afd78a

Observation 7212905d-590a-4677-bc73-d486d3d50ec8 · outbound

This paper cites Language models scale reliably with over-training and on downstream tasks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language models scale reliably with over-training and on downstream tasks

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.752393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.752393Z digest=sha256:acd436bd797eada5a67f49d4fb904c69e215b3ef47d23d33547b664ab17cec43

Observation f1e3447c-cb80-4840-93aa-5ed3ab6d4741 · outbound

This paper cites Broken Neural Scaling Laws.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Broken Neural Scaling Laws

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.756787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.756787Z digest=sha256:07078fd1cf37be3ea578b1b29cdaaa27d0dce43630eed3d2578e6bbd69c87997

Observation 76fd424c-0c4b-41c9-8969-76a406233349 · outbound

This paper cites Scaling laws for downstream task performance in machine translation,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling laws for downstream task performance in machine translation,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.761555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.761555Z digest=sha256:9bb10fbf6a434ee0e0e8d26416bd8de5cc6dc632c179cf8698025db7f4fe22ae

Observation 24a0ee2a-be06-4e0c-ab17-dac540d80973 · outbound

This paper cites Exploring the Limits of Large Scale Pre-training.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Exploring the Limits of Large Scale Pre-training

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.765677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.765677Z digest=sha256:643485f34c762442320234bf69a51411b813c42abee94505b49b306dc7996de5

Observation 0065568d-23ea-46ab-955f-c9b2bee4b5c9 · outbound

This paper cites Scaling Laws Do Not Scale.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws Do Not Scale

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.769818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.769818Z digest=sha256:06c560abd51a180e5dcde6c4492eed32f8fbcc0444ae998fe9360322be87688c

Observation 663ecb2d-e843-4e8e-8586-d528424257a7 · outbound

This paper cites Not-just-scaling laws: Towards a better understanding of the downstream impact of language model design decisions,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Not-just-scaling laws: Towards a better understanding of the downstream impact of language model design decisions,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.773993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.773993Z digest=sha256:82898f4ad2cb1b88caca7828948dd75db52d68f900b9b47fb969c68174cf091b

Observation 798b2385-8575-4476-9590-01cbf7243f93 · outbound

This paper cites Same Pre-training Loss, Better Downstream: Implicit Bias Matters for Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Same Pre-training Loss, Better Downstream: Implicit Bias Matters for Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.778229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.778229Z digest=sha256:4c9c9c73336c905e513facb3e8d8a8b2fcd7aa5de416787d37fee35f8d27482c

Observation 29406675-cb2b-4863-b3dd-2c3dd95f3e92 · outbound

This paper cites Overtrained Language Models Are Harder to Fine-Tune.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Overtrained Language Models Are Harder to Fine-Tune

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.782404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.782404Z digest=sha256:010e548d97570dee2fa97388f683dc26cc28111428e95790cf48f29d04c39b8a

Observation 4638aa3f-b743-452f-b280-1e9c4233a83c · outbound

This paper cites Rethinking fine-tuning when scaling test-time compute: Limiting confidence improves mathematical reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Rethinking fine-tuning when scaling test-time compute: Limiting confidence improves mathematical reasoning,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.786753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.786753Z digest=sha256:8378db00108a5eeb5c888e1e8e6fbd522741df606965665aeada5d3834beda15

Observation e9dd4514-ab59-4454-848f-e22edde220ed · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.791105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.791105Z digest=sha256:7d5c16efc81879f4db77915c50a26df2d191d2befe4ebe50f0a1de1a02f720db

Observation 33d1bc75-12dc-4f7f-984d-dfbfd9f856a9 · outbound

This paper cites BBQ: A Hand-Built Bias Benchmark for Question Answering.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search BBQ: A Hand-Built Bias Benchmark for Question Answering

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.795184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.795184Z digest=sha256:89fb8e0cc44a5b22e37f02e1320e55c258a51fd85b8542255e8047a380d3a021

Observation 50259960-870e-45ad-b591-0b717961a115 · outbound

This paper cites Inverse scaling can become U-shaped.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Inverse scaling can become U-shaped

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.799434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.799434Z digest=sha256:802cba2dd5d3b556ce2cd251339677a4acf79667606791e88909a437f814762c

Observation b52716c6-d685-48ae-a9a6-3271f62f6f1d · outbound

This paper cites s1: Simple test-time scaling.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search s1: Simple test-time scaling

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.803725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.803725Z digest=sha256:24bcc59162a6170e79586b4c4ffaac5d3243da5e61d3c4a4871f9c8311cd5953

Observation 5ad86b0f-93bd-4328-8700-4bce37b0ea1c · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.807957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.807957Z digest=sha256:8ea7719c32099cfa835a8372995199fda45e33da0865d667f88e6141f509cc40

Observation 5cbeb825-46fa-49b6-809b-3a8ec04067d0 · outbound

This paper cites Large Language Models are Zero-Shot Reasoners.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Large Language Models are Zero-Shot Reasoners

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.812191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.812191Z digest=sha256:159787e8958819a5d1f2fdf62c5b7110872e700f0a0823f10d695ea6f1b17a5e

Observation 2160b047-c53e-483c-8455-5b6271a68dcb · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.815981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.815981Z digest=sha256:057f9063a9a95eb6ecb05584c23563fc867fcc1dd53dfe37427d95fd7781f420

Observation 6eb5c544-c2dc-4356-ae67-c3edea4480a2 · outbound

This paper cites Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.820330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.820330Z digest=sha256:314a555946d886a4465c8a99ffc598b4222553482edd7bd24e9c579cf3c38429

Observation 395dad7c-5bf2-4692-9418-3263c1e2e7c9 · outbound

This paper cites Sequence to sequence learning with neural networks: What a decade,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Sequence to sequence learning with neural networks: What a decade,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.824523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.824523Z digest=sha256:29a34f0258648e1986af69a558494b635a50d8a88670c5179cae7da6b89f839e

Observation 43d03769-8801-4aad-8193-17701a25af25 · outbound

This paper cites AI doom from an LLM-plateau-ist perspective,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search AI doom from an LLM-plateau-ist perspective,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.828440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.828440Z digest=sha256:8f24499ac249b025e82d7d6e68505ee8a1a2f4f3827928c2fd2e763c86c4fc2f

Observation dd5603ca-4f42-4b38-a44e-be59e9426aa4 · outbound

This paper cites The first wave of AI innovation is over. here’s what comes next,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The first wave of AI innovation is over. here’s what comes next,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.832462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.832462Z digest=sha256:ae82111b0e43b966af2a7c2a483038a8a17123e0b6b0d5538552101b5220831d

Observation 0a53bca0-df12-4823-90de-32112cda7561 · outbound

This paper cites AI won’t plateau — if we give it time to think,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search AI won’t plateau — if we give it time to think,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.836408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.836408Z digest=sha256:f3a275cffa0c182ccf4cfe038e753a46ab7e67b28110b70efe49517d42b25299

Observation 22f4293a-5b65-424e-ba50-474d1949f1d5 · outbound

This paper cites Scaling data-constrained language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling data-constrained language models,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.840145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.840145Z digest=sha256:6696daab7e4546ccdbe6277569b64adb17667c03e0da7a3c379d85458f79ed96

Observation 023ee840-7c37-4322-992c-200a388fecf4 · outbound

This paper cites Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.844389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.844389Z digest=sha256:80c28312bea1d78fa658abd9cdf6f2048a3ae0457d7c8fc65ff0630173a9ddd7

Observation c6de41cc-bec8-40c6-bcd7-1aface91964a · outbound

This paper cites The rising costs of training frontier AI models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The rising costs of training frontier AI models

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.848494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.848494Z digest=sha256:764c88427e76de7c9e2f352bb3bc9d047fe57ebd9ddc92f3fd968380e64f7565

Observation 2fbe18fe-d46e-40fe-be04-734db104bfbf · outbound

This paper cites Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.852434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.852434Z digest=sha256:e175834375447521d258f65fd327b391b400af1155b6ae16e241b86ebd317469

Observation 8a60c330-fc02-4424-aa94-100361efb22b · outbound

This paper cites Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.856507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.856507Z digest=sha256:68ea489f4c7b86f4170d80f33266030e9117908f65f75fa29c6f82e1e71ca81c

Observation caa6192b-6ee1-41e8-a587-77f1a904d156 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Efficient Streaming Language Models with Attention Sinks

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.860710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.860710Z digest=sha256:546676a2b92df47567cbc6b0fecba027a04353b1d3fdd2414062ae7b191a9c07

Observation 71ffee8b-c47a-4cfb-a28b-3ff54ef3c5ad · outbound

This paper cites Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.864758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.864758Z digest=sha256:e70e75cc5ea3a6133f633208022fc1299777a61ee3cf6e79debf0db7e4f0c989

Observation 489eb03e-4f65-440f-b1aa-4fd2ce6adefb · outbound

This paper cites The curse of CoT: On the limitations of chain-of-thought in in-context learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The curse of CoT: On the limitations of chain-of-thought in in-context learning,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.868843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.868843Z digest=sha256:e12806fe68b683976c407955bc18c877ffcd8532470082438cff4c03936a7bca

Observation 634eb7db-c73e-4d8a-92e9-9d8456bc662d · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.872688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.872688Z digest=sha256:b454c423cddb838de38194648c5aac35fe8285e05aae00780e068eb5781d33fe

Observation 79cc6bdc-7dea-481c-bbe1-42664f1d2e42 · outbound

This paper cites Let's Verify Step by Step.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Let's Verify Step by Step

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.876992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.876992Z digest=sha256:515d88b665568d45aa24aeb4f3cf8024b0bf4f4716296d178111cec16d48525f

Observation a884701b-03ed-4fe4-9a68-c650ccb5ee85 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training Verifiers to Solve Math Word Problems

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.880977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.880977Z digest=sha256:f5d816d0d43c84bde4f0470268b2636a294834c5056e459de27c3820baa4cc9f

Observation e262f215-9b8c-4ac1-88db-db32dce362a4 · outbound

This paper cites When to solve, when to verify: Compute-optimal problem solving and generative verification for LLM reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search When to solve, when to verify: Compute-optimal problem solving and generative verification for LLM reasoning,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.884774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.884774Z digest=sha256:880ae690d5e381716782df91df564c8daedec8bdf49895d4e8c4d170fec9f5d6

Observation f052279f-279d-41c7-83a1-2efabc1fa6e6 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.888499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.888499Z digest=sha256:3008e27b4e33187a493490f61722579629091d0d99ea94b2bbcbf5a535c54eaf

Observation bf2247ff-05b5-4065-9db4-a24a122bd5a1 · outbound

This paper cites Solving quantitative reasoning problems with language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Solving quantitative reasoning problems with language models,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.892256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.892256Z digest=sha256:b9ef8a1805a387b85b3edf737b917821dc6a15aac099b72db34bedf0eba5ffb6

Observation 055e7f58-e5b6-429d-bf48-6546b15f8b3d · outbound

This paper cites Self-Refine: Iterative Refinement with Self-Feedback.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Self-Refine: Iterative Refinement with Self-Feedback

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.895889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.895889Z digest=sha256:79e823e3b6738268473b9876d47af4b9f0b4ba4e5d203c018ae1e0a8ee50bca6

Observation bf4a5fc3-212a-4143-b14f-e9186f1fb833 · outbound

This paper cites Cost-of-Pass: An economic framework for evaluating language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Cost-of-Pass: An economic framework for evaluating language models,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.899967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.899967Z digest=sha256:d58468dd330ddfa47361fded9196922edaae9b53ab940a225790e61a1c46fa1c

Observation 2aa08e6c-7e5a-4790-be4b-c704137801f3 · outbound

This paper cites Scaling Laws for Transfer.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Transfer

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.903603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.903603Z digest=sha256:eaa1c1b038f566af2e608e91fee09063b4728d2713b51fe22ec8b242d900c41b

Observation 45d542b2-0c07-4f17-ba0f-eb4f23d3ebc4 · outbound

This paper cites Reproducible scaling laws for contrastive language-image learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Reproducible scaling laws for contrastive language-image learning,

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.908059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.908059Z digest=sha256:0f5849af82ac7d5f213af044bbd2560b3e4aeb722af112412cad9c18395d0c05

Pith citing papers

Observation 5e5e584d-abf1-49fc-9722-19276522f94e · inbound

Two AI Metrics Diverged: Will it Make All the Difference? cites this paper.

Two AI Metrics Diverged: Will it Make All the Difference? A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:56.137222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-02T12:29:24.439779Z digest=sha256:f6dfc614b768b9b7d98e90d80069d6ef7b2e7bdee37a5340bf9d6b60bfd76b00